Skip to content

PubChem PUG REST

PubChem PUG REST is an HTTP API for retrieving selected chemical records, properties and assay information, with documented identifier inputs, structure searches and operation-specific output formats.

Catalog updated ·

Overview

PubChem PUG REST provides programmatic access to selected PubChem information through HTTP requests. It supports focused retrieval in scripts, web applications and third-party tools, rather than requiring an installable client package. The official specification describes URLs organized into input, operation and output components. This entry concerns the hosted API itself, not a wrapper or third-party MCP server.

Compound inputs include CID, name, SMILES, InChI, InChIKey and molecular formula, alongside structure-search inputs. Substance and assay domains are also documented. Available operations include retrieving full records, synonyms, identifiers, compound properties and assay summaries, plus similarity and identity searches. Outputs depend on the operation and can include JSON, XML, SDF, PNG, TXT or CSV. These choices allow a workflow to request a specific property table or structure record instead of treating every request as an identical record export.

The tutorial positions PUG REST for focused requests, not millions of individual calls. Clients should respect published request limits, account for dynamic throttling and use bounded retries; bulk retrieval requires suitable bulk-download workflows. Special URL characters need encoding, while InChI and SDF inputs use POST. The programmatic-access overview distinguishes its short synchronous requests from PUG View’s complete summary reports and third-party annotations. Retrieving data does not experimentally validate bioactivity or establish reuse rights for every annotation. The supplied evidence establishes neither a service code repository nor a code license.

Key Features

  • HTTP request paths separate the input specification, requested operation and output format.
  • Compound queries accept CID, name, SMILES, InChI, InChIKey and formula inputs; substance and assay domains are also documented.
  • Retrieval operations cover full records, synonyms, identifiers, compound properties and assay summaries.
  • Structure-oriented operations include similarity and identity searches.
  • Operation-dependent outputs include JSON, XML, SDF, PNG, TXT and CSV; InChI and SDF inputs use POST.
  • Focused synchronous access is documented with request limits and dynamic throttling, rather than unrestricted bulk per-record retrieval.

Use Cases

  • Suggested application: enrich a curated compound list with documented properties while preserving the PubChem identifier and retrieval date for each result.
  • Suggested application: obtain SDF structure records for downstream inspection or cheminformatics processing, evaluating their suitability separately.
  • Suggested application: retrieve synonyms and identifiers to support a chemical-name reconciliation workflow; assess ambiguous matches before accepting them.
  • Intended evaluation: assess whether documented similarity searches or assay summaries meet a research workflow’s information needs without treating retrieved results as experimental validation.

How to Use

  1. Consult the PUG REST specification and choose the appropriate domain, input identifier, operation and output format. Prefer a stable record identifier when available.
  2. Inspect the documented aspirin SDF example to understand a structure-record request before adapting it to your chosen compound.
  3. For focused property retrieval, examine the MolecularFormula and InChIKey JSON example. Confirm that the returned fields meet your downstream requirements.
  4. Follow the specification’s input rules: encode special URL characters and use POST for InChI or SDF inputs. Inspect HTTP status and returned data, retaining the source identifier and retrieval date.
  5. Apply rate limiting and bounded retries using the official tutorial. Plan bulk retrieval separately rather than issuing unrestricted parallel per-record requests.
  6. Review the programmatic-access overview when choosing between focused PUG REST retrieval and PUG View reports. Evaluate data interpretation and annotation reuse requirements separately from API access.

Related resources

PubChem MCP (cyanheads) connects MCP clients to PubChem compound and bioassay data, with identifier and structure searches, property retrieval, safety records, cross-references and 3D conformers.

Open sourceTypeScript

Cheminformatics · Scientific Data

PubChem MCP (PhelanShao) is a Python MCP server that lets AI clients retrieve PubChem compound properties and structures by name or CID, with JSON, CSV, XYZ and downloadable structure-file outputs.

Open sourcePython

Cheminformatics · Scientific Data

PubChemPy

Open Source

PubChemPy is a Python wrapper for the PubChem PUG REST API, supporting chemical searches, compound-property retrieval, standardization, file-format conversion and depiction.

Open sourcePython

Cheminformatics · Scientific Data

Scientific Agent Skills provides procedural guidance for AI agents working with scientific packages, databases and research workflows, including cheminformatics, spectroscopy and materials analysis.

Open sourcePython

Cheminformatics · Scientific Data

Benchling MCP (longevity-genie) is a Python MCP server that connects AI clients to Benchling notebook entries, biological sequences, projects, and entity search using API credentials.

Open sourcePython

Lab Automation · Scientific Data

ChEMBL MCP (cyanheads) connects MCP clients to ChEMBL compound, target, bioactivity and drug records, with structure searches and optional DuckDB analysis of larger activity sets.

Open sourceTypeScript

Drug Discovery · Scientific Data

Works with

Used in Recipes

Official

Chemistry Research with Claude + PubChem

Resolve chemical names to PubChem identifiers and retrieve traceable compound properties from Claude Desktop through a locally installed PubChem MCP server.

PubChem MCP (cyanheads) + PubChem PUG REST

Analyze Molecules

Level: Intermediate Cost: Mixed Privacy: Cloud ~30 min

Claude Desktop

View Setup →

Related guides

Workflows

AI for Drug Discovery: Agents, MCP Servers and Platforms

Choose tools by workflow stage: target evidence, protein structures, known binding sites, molecular generation, property analysis and screening. Compare research agents, MCP integrations, hosted APIs and platforms, with a proposed end-to-end example and practical evaluation gates.

Overview

MCP vs Skills vs Agents vs APIs: What's the Difference?

Understand how APIs, MCP servers, Skills, and agents divide interface, procedural guidance, execution, and planning responsibilities in chemistry workflows—and how to choose or combine them without confusing access with scientific validity.

Workflows

Building an AI Cheminformatics Workflow with RDKit

Design a traceable RDKit workflow for molecular validation, standardization, descriptors, fingerprints and chemical searches. Compare Python, MCP and Skill access, then add bounded PubChem enrichment without confusing tool execution with scientific validation.