AI News

OpenAI vs Grok Bot and Jev: What DevDay Actually Shipped

OpenAI announced more than 20 products and changes at DevDay 2026. NetworkChuck's live reaction asks a narrower question: if you already run agents and use fast decision models, what is genuinely new for you? He sees Dots alongside Grok Bot and his own Hermes setup, and the Decisions API alongside TypeSafe's Jev. That makes for a useful comparison, provided we do not turn a livestream reaction into a claim that any company copied another or that an unreleased API has won a benchmark.

The short answer: Dots promises a managed, persistent agent inside ChatGPT; Grok Bot offers a team of always-on agents; a custom Hermes setup gives its builder more direct control. Decisions API and Jev both address small, bounded decisions inside software, but access to each was limited when the stream aired: Jev was in early access and Decisions API in limited preview. NetworkChuck had not personally tested Dots, Grok Bot, or GPT-6.1 Sol in this livestream.

Watch NetworkChuck's DevDay Livestream

Credit: NetworkChuck's livestream, published 29 September 2026. He says the stream is not sponsored by OpenAI. This article uses the supplied transcript and checks launch status against the linked first-party sources as of 1 October 2026. The stream also includes extended audience Q&A; the chapter links below point to the product discussion.

Dots, Grok Bot, and the Agent You Already Run

At 03:00, NetworkChuck opens OpenAI's announcement and immediately compares Dots with Grok Bot and his Hermes agents. These products overlap in the job they pitch: maintain context, work in the background, use a computer and connected tools, and bring important decisions back to a person. But overlap in the pitch does not establish who built a feature first, which agent finishes more tasks, or whether moving systems is worth it.

OptionWhat the source describesThe trade-off to test
OpenAI DotsOne primary dot at launch, with its own cloud computer, ChatGPT context, connected apps, optional local-computer access, and Work or Codex handoffs.Convenient inside OpenAI's ecosystem, but access, deeper-work allowance, permissions, and regional rollout matter.
Grok BotMultiple bots can coordinate, use their own cloud computer, and work across apps; xAI lists a beta and separate Bot usage.A multi-bot model may fit teams, but test the actual workflow and approvals, not the number of agents in a demo.
NetworkChuck's Hermes setupHe describes agents in Slack and on his own machines, with a choice of models, local integrations, and personally managed memory.More control and portability for a technical owner; more setup, security work, and maintenance are yours.

NetworkChuck says he had not yet received Dots access and had not used Grok Bot. His preference for a configurable agent is a useful point of view, not the result of a side-by-side test. For readers in Portugal, OpenAI's initial personal Pro Dots rollout excludes the EEA; Business Premium is listed across supported regions, and Enterprise beta requires an administrator to enable it. Dots also starts with one primary agent, with specialist or multiple-dot setups in pilots or future plans.

Decisions API and Jev: Same Job, Different Evidence

At 31:43, the livestream turns to Jev: a model TypeSafe built to return structured choices, scores, and probabilities for software to act on, rather than long prose. The example is a YouTube comment. Is it spam, praise, or an attempt to redirect viewers? The surrounding application uses those answers to route the comment. This is a much narrower problem than having an agent write or operate a whole project.

OpenAI's Decisions API similarly takes a developer-defined set of questions with finite answers and uses Luna's intelligence to classify or choose a next action. OpenAI says its context can include text or images. It was announced in limited preview, with broader release planned; that is not the same as production availability for every builder. TypeSafe describes Jev as an early-access System One model with typed probabilistic outputs and publishes caveats around its speed and cost comparisons. Ollama 0.35.0 also added a local /v1/systemone interface based on Jev's API, with Nimble and Tev1 listed as available models. That is an interoperability option, not a claim that Jev itself runs locally through Ollama.

NetworkChuck's “coming for Jev” is his interpretation of a shared use case, not proof of copying or a settled winner. An independent Every preview test found the winner varied by task: Decisions did better on a small computer-control replay, while Jev was faster in a conversation-classification test. Neither sample establishes broad reliability or a stable cost advantage; OpenAI had not announced Decisions pricing in that comparison.

GPT-6.1 Sol: A Candidate, Not Chuck's Verdict

At 19:35, NetworkChuck reacts to GPT-6.1 Sol. OpenAI positions it near Astra on selected coding, computer-use, and professional-work evaluations at one-fifth of Astra's standard input and output token prices. That is a vendor comparison of model prices and published tests, not a promise that every completed agent job costs one-fifth as much. Retries, tool use, reasoning settings, and review time can change the total.

Chuck wants to try Sol as the main model behind some of his Hermes agents. He is candid that he had not benchmarked 6.1 Sol when he spoke. His disappointment with GPT-6 Sol and enthusiasm for Astra are personal usage reports; the new model needs its own workload test before it earns a role in his stack.

Pro 500 and Ultrafast: When Speed Is Worth Paying For

The 40:23 segment covers the new $500/month Pro 500 plan. OpenAI lists the highest included Pro usage and Astra Ultrafast access for that tier; Pro 100 and Pro 200 do not include Ultrafast. The DevDay recap describes up to 8x faster token generation in Codex and up to 6x via the API. Those are speed claims for token generation, not a guarantee that a research or coding job finishes 8x faster end to end.

NetworkChuck says he might buy the higher tier but does not yet know if Ultrafast is worth it for his work. That is the right buying question. Measure completed tasks, total wall time, corrections, and how quickly the allowance is consumed. OpenAI's API Ultrafast guide is a separate pricing and rate-limit surface from a personal ChatGPT subscription.

The Other Changes He Paused On

In the 45:32 lightning round, he is excited by voice in the refreshed Codex CLI, less interested in cloud Codex because his own remote setup already works, and curious about ChatGPT Space. These are preferences shaped by his existing tools. For a team without his infrastructure, reusable Codex cloud environments and shared settings may be the more relevant change.

He also flags Agents API computer use and Sign in with ChatGPT. The latter separates identity from eligible plan-backed usage in participating apps; signing in does not give every third-party tool unlimited access to a user's subscription. For the full official launch inventory, including plugins and Space, use our DevDay keynote guide.

What Was Not Broadly Available at Launch

Those limits are more actionable than speculation about an unannounced device or who will win the agent market. This was a live reaction to a launch, not a performance audit of every new product.

A Test for Your Own Agent Stack

  1. Pick one repeated task. For example, triage 50 anonymized comments or prepare one weekly support summary.
  2. Write the rules before choosing a vendor. Define allowed actions, a human approval point, and what counts as a correct result.
  3. Run the same sample through available tools. Compare your current agent with Dots or Grok Bot only where you actually have access. Compare decision models only after Decisions becomes available to you.
  4. Score finished work. Record accuracy, time, total cost, failed handoffs, and how much work a person had to redo. Keep the raw examples for a later retest.

That test may reveal a new product is better for you. It may also show that the setup you already trust is sufficient. Both results are useful.

Turn This Into a Business Idea

The most interesting business may not be another general assistant. It may be a narrow, repeated job for a customer you already understand. This prompt uses the agent, decision, and model-routing ideas above to find a testable offer. It works in any capable assistant and does not require DevDay preview access.

Business idea prompt

Find a paid problem hiding in these launches

Start from your own skills and customer access, then pressure-test three ideas before building one.

Ready to copy

Livestream Chapters

TimeTopicTimeTopic
00:14Livestream opening03:00Dots first reaction
06:35Space and specialist dots17:12Dots, Grok Bot, and Hermes
19:35GPT-6.1 Sol31:43Decisions API and Jev
40:23Pro 500 and Ultrafast45:32Codex, Agents API, and sign-in

Sources and Link Map

Common questions

Did NetworkChuck test Dots, Grok Bot, or GPT-6.1 Sol head to head?
No. In this livestream he says he has not yet received Dots access, has not used Grok Bot, and has not benchmarked GPT-6.1 Sol. His comparisons are informed reactions, not a controlled product test.
Is OpenAI Decisions API already generally available?
No. OpenAI announced a limited preview at DevDay and planned a broader release in the following days. Check current access before building a production dependency.
Are Decisions API and Jev the same model?
No. Both target bounded classification and routing decisions, but OpenAI says Decisions accepts text and image context, while TypeSafe describes Jev as a System One model with typed probabilistic outputs. A shared use case is not proof of identical architecture, pricing, quality, or reliability.
Does Pro 500 make Astra Ultrafast unlimited?
No. It has the highest included Pro usage and Ultrafast access, but allowances and optional credits still apply. OpenAI lists Ultrafast only on Pro 500 among personal Pro tiers.
Share
X LinkedIn Reddit
Build Yours

Want a system
like this one?

Book a free 30-minute call. We map your situation, identify the highest-impact automation, and figure out if we are a fit.

Book Free 30-min Call