Can GPT-6 Astra Run Agents? What We Found Testing It Live
OpenAI launched GPT-6 Astra on computer use and browsing, but on Chat Completions it has no function calling. The two 400 errors we hit on September 8, 2026, and which models to use for agents instead.
TL;DR: Yes, but only through OpenAI's Responses API. On Chat Completions, GPT-6 Astra has no function calling: when we tested it live on September 8, 2026, every tool-bearing request we tried returned a 400. Use Astra for chat, vision and deep reasoning. For agents on TulexAI, pick a tool-capable model like Claude Fable 5.1.
Can GPT-6 Astra run agents?
Yes, with one condition: the endpoint. OpenAI released GPT-6 Astra in limited preview on September 3, 2026 and made it generally available on September 4 (OpenAI's announcement). OpenAI's reasoning guide is blunt about tools:
"Chat Completions does not support function calling with GPT-6 Astra."
Function calling, web search, file search, code interpreter and computer use for Astra are all Responses API only. If your agent talks to OpenAI through the Responses API, Astra can call tools. If it speaks Chat Completions, as TulexAI's API does, it cannot.
What happened when we sent tools to Astra over Chat Completions?
We tested Astra live with our own API key on September 8, 2026. Text, streaming and image input returned 200, so those went live that day. Then we tried function tools, the one thing every agent needs.
| Request to Chat Completions | Result we observed on 2026-09-08 |
|---|---|
| Text, streaming or image input, no tools | 200 |
| Tools plus any reasoning effort | 400 |
Tools plus reasoning_effort set to 'none' | 400 |
The first 400, verbatim as we saw it that day (trimmed where marked with "..."):
"Function tools with reasoning_effort are not supported for gpt-6-astra ... set reasoning_effort to 'none'"
So we did what it asked. The second 400, also verbatim:
"'reasoning_effort' does not support 'none' with this model. Supported values are: 'low', 'medium', 'high', and 'xhigh'."
The first error tells you to use 'none'. The second says 'none' is not allowed. No request shape we tried on Chat Completions got a tool call through, which matches OpenAI's own sentence above.
Why does this surprise people who read the launch news?
Because agent work was the headline. OpenAI calls Astra "the most intelligent and aligned model in the world" and says it is state of the art for computer use, browsing, software engineering, cybersecurity, science and professional work. Those are vendor claims, not independent measurements. CNBC's launch report describes Astra carrying out computer tasks such as filling out online forms, updating customer records and organizing calendars.
Every one of those demos depends on tools, and for Astra those tools sit behind the Responses API. Point a Chat Completions agent at gpt-6-astra and you get a 400 instead of a filled-in form. The model can do the work. The endpoint won't let it.
What does this mean if you use GPT-6 Astra on TulexAI?
TulexAI calls OpenAI's Chat Completions endpoint, so we inherit the limit. We don't add tools to Astra. Here is what you get:
- Works: chat, streaming, image input and long-form reasoning, in the web app and over our OpenAI-compatible /v1 API.
- Refused: any request that carries tools, with a clear error that says what went wrong instead of OpenAI's raw 400.
- Cost and plans: 25 Platform Tokens per 750 tokens of prompt plus reply (reasoning included, 25 PT minimum per request), on Pro ($23/mo), VIP and Elite. Not on Basic.
- Output cap: replies are capped at 8,192 output tokens on TulexAI, below the 128,000 max output in OpenAI's model docs.
Our agent setup docs say the same: Astra is for chat, vision and long-form reasoning, and agent work belongs on a tool-capable model. Full specs are on the GPT-6 Astra model page.
Which models should you use for agents on TulexAI instead?
| Model | Function tools on TulexAI | Plans | Why pick it |
|---|---|---|---|
| Claude Fable 5.1 | Yes | Pro ($23/mo) and up, 25 PT per 750 tokens | Ties Astra on Intelligence Index v4.3 (53) and the Coding Agent Index (62) |
| Claude Opus 5 | Yes | Every plan, from Basic ($11/mo) | Tool-capable Claude without the Pro gate |
| GPT-5.6 Sol | Yes | Every plan | Stay with OpenAI and keep working function calls |
Fable 5.1 is the closest swap. Anthropic released it on September 1, 2026 (Anthropic's announcement). On Artificial Analysis' Intelligence Index v4.3 it scores 53 at max effort with fallback, level with Astra at max, and the two tie again at 62 on the Coding Agent Index.
One workable split: let a tool-capable model drive the agent loop, and take the hardest no-tools reasoning question to Astra in chat. More context in our GPT-5.6 Sol, Luna and Terra comparison and Claude Fable 5 vs Kimi K3 vs GPT-5.6 Sol.
Where does GPT-6 Astra shine when you don't need tools?
Every figure below is an independent measurement from Artificial Analysis (Intelligence Index v4.3 write-up, published September 7, 2026, and the Astra vs Fable 5.1 comparison). Caveat: Artificial Analysis ran Astra at max or high reasoning effort. TulexAI runs Astra at OpenAI's default effort, which OpenAI does not document. Read these as the model's ceiling, not the exact scores you'll get in our chat.
| Artificial Analysis metric | GPT-6 Astra | Claude Fable 5.1 | GPT-5.6 Sol |
|---|---|---|---|
| Intelligence Index v4.3 | 53 at max (#3 of 200), 51 at high | 53 (max with fallback) | 47 (max) |
| Coding Agent Index | 62 | 62 | - |
| Output tokens per Index task | ~27k | ~78k | - |
| Cost per Index task | $3.26 | $7.63 | - |
| AutomationBench-AA | 68% | 59% | - |
| Terminal-Bench v4.0 | 59% | 52% | - |
| Hallucination rate (max effort) | 51% | - | 92% |
- Efficiency at the top. Same 53 as Fable 5.1 with about a third of the output tokens, and $3.26 versus $7.63 per task. Astra's API price is $10 input / $50 output per 1M tokens (OpenAI pricing).
- Fewer made-up answers than Sol. At max effort the hallucination rate drops from GPT-5.6 Sol's 92% to 51%. Better, still far from zero, so check anything that matters.
- Big inputs. A 1,050,000-token context with text and image input, knowledge cutoff April 30, 2026.
- It thinks first. 334 seconds to first token at max effort, thinking included, and 37 seconds at high. Send it the hard question, not the quick rewrite.
The irony: Astra beats Fable 5.1 on AutomationBench-AA and Terminal-Bench v4.0, exactly the work you'd wire tools into. Our read: over Chat Completions, Astra is a superb reasoner you talk to, not an agent you deploy.
Create your TulexAI account and choose Pro ($23/mo) to try GPT-6 Astra in chat, with Claude Fable 5.1 and GPT-5.6 Sol in the same dropdown for work that needs tools. Compare plans on the pricing page, or see how it stacks up against a single ChatGPT plan on our ChatGPT alternative page.
Frequently Asked Questions
Does GPT-6 Astra support function calling?
Only through OpenAI's Responses API. OpenAI's reasoning guide states: "Chat Completions does not support function calling with GPT-6 Astra."
Why does GPT-6 Astra return a 400 when I send tools?
On Chat Completions it's a catch-22. In our live test on September 8, 2026, tools with any reasoning effort returned a 400 asking for reasoning_effort 'none', and tools with 'none' returned a 400 listing 'low', 'medium', 'high' and 'xhigh' as the supported values.
Can I use GPT-6 Astra with my coding agent on TulexAI?
Not for agent work. TulexAI calls Chat Completions, where Astra has no function tools, so a tool-bearing request is refused with a clear error. Give your agent Claude Fable 5.1, Claude Opus 5 or GPT-5.6 Sol instead.
Will I get Astra's benchmark scores on TulexAI?
Don't assume so. Artificial Analysis measured Astra at max effort (53) and high effort (51) on Intelligence Index v4.3. TulexAI runs Astra at OpenAI's default effort, which OpenAI does not document, so results may differ.
Which TulexAI plans include GPT-6 Astra?
Pro ($23/mo), VIP and Elite, not Basic. Astra costs 25 Platform Tokens for every 750 tokens of prompt plus reply, reasoning included, with a 25 PT minimum per request, and replies are capped at 8,192 output tokens. Create your account to try it in chat.
Ready to consolidate your AI tools?
40+ AI models - GPT-5.6, Claude Opus 5, Gemini, Flux, Sora & more. One subscription from $11/mo.
Try 1 Free PromptContinue reading
What Is Claude Opus 5.5? What Changed, What It Costs, and Where It Falls Down
Anthropic's Claude Opus 5.5 costs 20% less per token than Opus 5, and Anthropic says it performs at the level of Claude Fable 5.1 on most work. What changed, what it costs on the API and on TulexAI, where it falls down, and every way to get it.
What Is GPT-6 Astra? OpenAI's New Flagship, and How to Actually Get It
OpenAI's GPT-6 Astra ties Claude Fable 5.1 on the neutral scoreboard, but ChatGPT Plus only gets it in Work and Codex. What it is, every way to access it, what it costs, and the catches.