• Nextool AI
  • Posts
  • OpenAI and Cerebras Make GPT-5.6 Sol 7× Faster

OpenAI and Cerebras Make GPT-5.6 Sol 7× Faster

PLUS: IBM Partners With OpenAI to Bring GPT-5.6 to Enterprises

In partnership with

OpenAI is making moves on both speed and enterprise adoption, while Microsoft is simplifying its AI strategy. Cerebras is powering GPT-5.6 Sol at up to 750 tokens per second, IBM is bringing OpenAI’s models deeper into global enterprises, and Microsoft is cutting underused Copilot features while merging its consumer and business apps.

In today’s post:

  • GPT-5.6 Sol can now hit 750 tokens per second

  • IBM just gave OpenAI an enterprise shortcut

  • Microsoft is killing Copilot features

SPONSORED BY

Blu Dot surpasses 2,000% ROAS with self-serve CTV ads

Home furniture brand Blu Dot blew up on CTV with help from Roku Ads Manager. Here’s how:

After a test campaign reached 211,000 households and achieved 1,010% ROAS, the brand went all in to promote its annual sales event. It removed age and income constraints to expand reach and shifted budget to custom audiences and retargeting, where intent was strongest.

The results speak for themselves. As Blu Dot increased their investment by 10x, ROAS jumped to 2,308% and more page-view conversions surpassed 50,000.

“For CTV campaigns, Roku has been a top performer,” said Claire Folkestad, Paid Media Strategist, Blu Dot. “Comping to our other platforms, we have seen really strong ROAS… and highly efficient CPMs, lower than any other CTV partner we've worked with.”

Using Roku Ads Manager, the campaign moved from a pilot to a permanent performance engine for the brand.

What’s Trending Today

PARTNERSHIP

OpenAI and Cerebras are making frontier AI feel instant

Image Credits: Cerebras

Cerebras and OpenAI just unveiled Ultrafast Mode. It runs GPT-5.6 Sol at remarkable speeds. The headline number is 750 output tokens per second. But raw speed isn't the interesting part. It's what happens when waiting disappears from AI workflows.

  • GPT-5.6 Sol Ultrafast reaches 750 output tokens per second.

  • Cerebras says it runs 11x faster than Fable 5.

  • It completed Humanity's Last Exam in 11 hours.

  • Fable 5 took over 78 hours in Cerebras' comparison.

  • GDP-Val workloads saw a 5.6x end-to-end speedup.

  • Cerebras uses wafer-scale chips with 44 GB of SRAM.

  • Ultrafast launches through OpenAI's API in limited preview.

AI speed isn't just about getting answers sooner. Latency changes how people actually use these systems. Minutes encourage you to leave and do something else. Seconds let AI remain inside your train of thought. That matters even more as agents handle longer tasks. The next frontier may be intelligence without the waiting.

BREAKTHROUGH

IBM is turning thousands of consultants into OpenAI specialists

Image Credits: IBM Newsroom

IBM just partnered with OpenAI. But this deal is bigger than another AI integration. It gives OpenAI direct access to major enterprises. It gives IBM another frontier AI provider. And it shows where the AI battle is moving next.

  • IBM will create a dedicated OpenAI practice inside IBM Consulting.

  • Tens of thousands of consultants will train on OpenAI technology.

  • Training will cover Codex, APIs, cybersecurity, and consulting solutions.

  • GPT-5.6, Codex, and ChatGPT Work will enter IBM’s consulting platform.

  • IBM will build solutions for finance, government, telecom, and retail.

  • The deal strengthens IBM’s model-agnostic approach alongside its Granite models.

  • For OpenAI, IBM becomes another distribution channel into global enterprises.

The AI race is becoming a distribution race. Great models matter, but deployment matters just as much. Enterprises rarely adopt important technology with one API call. They need consultants, integrations, governance, training, and support. IBM already has those relationships. That could make this partnership more important than it looks.

RESEARCH

Microsoft is cutting AI features to make Copilot simpler

Image Credits: Microsoft

Microsoft spent years adding features to Copilot. Now, it is starting to remove them. The company is merging its consumer and business Copilot apps. Several experiments are disappearing too. The bigger lesson is about product focus.

  • Microsoft will merge Copilot and Microsoft 365 Copilot functionality.

  • Group Chats and AI-generated podcasts will disappear.

  • Copilot Labs features are also being removed.

  • Deep Research will vanish for consumers by August 18.

  • Microsoft is retiring Mico, Copilot’s animated assistant character.

  • Generated files will move from Copilot into OneDrive.

  • Microsoft says the goal is one simpler, cohesive experience.

AI products are entering a different phase. Adding more features no longer guarantees a better product. Too many tools can make the experience harder. Users usually want one clear place to get things done. Microsoft seems to be learning that now. In AI, simplicity may become a stronger advantage than novelty.

Free Guides

My Free Guides to Download:

🚀 Founders & AI Builders, Listen up!

If you’ve built an AI tool, here’s an opportunity to gain serious visibility.

Nextool AI is a leading tools aggregator that offers:

  • 500k+ page views and a rapidly growing audience.

  • Exposure to developers, entrepreneurs, and tech enthusiasts actively searching for innovative tools.

  • A spot in a curated list of cutting-edge AI tools, trusted by the community.

  • Increased traffic, users, and brand recognition for your tool.

Take the next step to grow your tool’s reach and impact.

That's a wrap:

Please let us know how was this newsletter:

Login or Subscribe to participate in polls.

Reach 150,000+ READERS:

Expand your reach and boost your brand’s visibility!

Partner with Nextool AI to showcase your product or service to 140,000+ engaged subscribers, including entrepreneurs, tech enthusiasts, developers, and industry leaders.

Ready to make an impact? Visit our sponsorship website to explore sponsorship opportunities and learn more!