Introducing open source Cambrian Core. Explore
PARALLEL RESPONSES API

OpenAI-compatible answers grounded in the live web

Stream cited factual answers with zero prompt engineering or fragile search tool loops. Save up to 80% of agent context window tokens.

Responses Quickstart$2.50 / 1K queries
from parallel import Parallel

client = Parallel(api_key="PARALLEL_API_KEY")

# Drop-in synthesized research with OpenAI compatibility
response = client.responses.create(
    question="Compare Apple M4 Max memory bandwidth and unified architecture to Nvidia RTX 5090",
    stream=True
)

for chunk in response:
    print(chunk.delta, end="", flush=True)

print("\n\nVerified Citations:")
for citation in response.citations:
    print(f"[{citation.id}] {citation.title} - {citation.url}")
Architecture

Why developers choose Responses API

Drop-in OpenAI Compatibility

Switch the base URL in your existing LangChain, LlamaIndex, or OpenAI SDK setup and get instant web research without code rewrites.

Context Window Optimization

Eliminate 30,000+ tokens of raw HTML search results. Responses API returns concise, cited prose consuming under 800 tokens.

Real-Time Streaming

First token latency in under 1.5 seconds with Server-Sent Events (SSE) streaming directly into your conversational chat UI.

Verifiable Inline Footnotes

Every factual claim includes bracketed footnote citations linking to primary source publisher pages with Basis confidence scores.

Institutional Grade

Built for enterprise, secure by design

SOC 2 Type II certified, GDPR compliant, and strict Zero Data Retention (ZDR) guarantees. Your confidential queries and internal data never touch model training.

SOC 2 Type II Certified
Zero Data Retention (ZDR)
Zero Model Training

FAQs

The Responses API is a drop-in endpoint that takes a natural language question, searches the live web, verifies sources, and returns a fully synthesized, cited answer in seconds without requiring multi-step tool loops in your LLM.

Where agents find answers

Start building autonomous workflows today with up to 5,000 free requests per month.