Responses API
# High-quality web research, in seconds
## Build fast and accurate web research experiences into your product.
Features
Fast, predictable, drop-in
High quality, low latency
Responses returns simple questions in just a couple seconds and can do deep research inside a minute. Perfect for conversational agents and user-facing UIs.
Flat, predictable cost
Each Responses API call costs the same, no matter how many tokens Parallel consumes under the hood. Scale usage confidently from 1 call to 1 million calls.
OpenAI compatible
Use Parallel's Responses API through OpenAI's Python or TypeScript SDKs. Switching to Parallel is a three-line diff.
Use Cases
In your product, or inside your agent
In-product web research
Build delightfully smart product experiences, powered by fast web research.
A subagent specialized for research
Let your agent delegate web research tasks to Parallel Responses, making your agent faster, smarter, and much more token-efficient.
Start building for free
Get started with our APIs in seconds. Run up to 5,000 requests per month for free.
Agent onboarding prompt:
Use curl to read parallel.ai/agents.md and perform the setup to install Parallel