Responses API
# High-quality web research, in seconds
## Build fast and accurate web research experiences into your product.
Features
Fast, predictable, drop-in
High quality, low latency
Responses returns simple questions in just a couple seconds and can do deep research inside a minute. Perfect for conversational agents and user-facing UIs.
Flat, predictable cost
Each Responses API call costs the same, no matter how many tokens Parallel consumes under the hood. Scale usage confidently from 1 call to 1 million calls.
OpenAI compatible
Use Parallel's Responses API through OpenAI's Python or TypeScript SDKs. Switching to Parallel is a three-line diff.
Use Cases
In your product, or inside your agent
In-product web research
Build delightfully smart product experiences, powered by fast web research.
A subagent specialized for research
Let your agent delegate web research tasks to Parallel Responses, making your agent faster, smarter, and much more token-efficient.
Start building for free
Get started with our APIs in seconds. Run up to 5,000 requests per month for free.
Agent onboarding prompt:
Use curl to read parallel.ai/agents.md and perform the setup to install Parallel
Start building for free
Get started with our APIs in seconds. Run up to 5,000 requests per month for free.
Use curl to read parallel.ai/agents.md and perform the setup to install Parallel

