- The AI Report
- Posts
- π OpenAI and Anthropic Price War + Claude's Invisible Watermarks + Groq Raises $350M
π OpenAI and Anthropic Price War + Claude's Invisible Watermarks + Groq Raises $350M
Plus: Google announces Gemini 3.7 flash, Nvidia Nemotron 3.5 Lightning now available, Gemini becomes Google's fastest-growing product, and Nvidia discloses $21bn stake in SpaceX
The AI landscape is shifting dramatically
The escalating price war between OpenAI and Anthropic highlights the urgent need for innovation as they face competition from cost-effective Chinese AI models.
Plus, Anthropic unveils its invisible watermarking technology while Groq secures $350M to pivot towards neocloud services.
Learn AI in 5 minutes a day
You don't have to scroll every AI thread, track every new tool, or watch every demo.
The Rundown AI breaks it all down for you β the latest AI news, tools, and tutorials in one free 5-minute email every morning.
Trusted by 2M+ professionals at Apple, Google, and NASA.
π OpenAI and Anthropic in price war as Chinese AI rivals gain ground

- 80% price cut - OpenAI has dramatically reduced the price of its GPT-5.6 Luna model from $1 to just $0.20 per million input tokens, making it significantly more competitive in the current market.
- Chinese competition rises - Chinese AI developers are gaining traction in the market, with companies like DoorDash and Airbnb beginning to adopt their models to manage rising costs, highlighting the increasing pressure on US labs.
- Shift to usage-based billing - US AI labs are transitioning some enterprise customers from flat-rate subscriptions to usage-based billing, which could lead to further cost scrutiny and shifts in customer loyalty.
The Bigger Picture: The escalating price war among US AI labs signifies a critical juncture in the industry, as OpenAI and Anthropic scramble to defend their market share against the encroaching threat of cost-effective Chinese models. This shift not only reflects the immediate pressures of rising operational costs but also indicates a potential long-term transformation in how AI services are priced and consumed. As companies increasingly gravitate towards usage-based billing, the dynamics of customer loyalty may shift, favoring models that deliver high performance at lower costs. The next 6-12 months will likely see intensified competition, with US firms needing to balance innovation and affordability to maintain their edge in a rapidly evolving landscape.
π Anthropic explains how Claude's invisible text watermarks will work

- Invisible watermarking - Claude's watermarking technique embeds undetectable patterns in text by manipulating low-stakes word choices, allowing for content verification without altering readability.
- Randomness with a purpose - Instead of relying on arbitrary random number generation, the watermarking process uses a specific key and preceding words to guide the model's word selection, ensuring a consistent watermark.
- Content authenticity - This innovation aims to address concerns about content authenticity and traceability, providing a mechanism for verifying the source of generated text without impacting user experience.
The Bigger Picture: Anthropic's introduction of invisible watermarking for Claude reflects a growing emphasis on content authenticity in the AI landscape. As generative models proliferate, the ability to trace and verify AI-generated content will become critical, especially in contexts where misinformation can have serious repercussions. This move not only positions Anthropic as a leader in responsible AI deployment but also sets a precedent for competitors to follow suit. In the next 6-12 months, we may see increased regulatory scrutiny and demand for similar technologies as the industry grapples with the implications of AI-generated content on trust and accountability.
π’ Groq raises $350M to fuel its pivot from AI chips to neocloud

- Valuation drop explained - Groq's valuation has decreased significantly after losing its founding team to Nvidia, but the company frames this as a revaluation post-licensing deal rather than a down round.
- Shift to neocloud services - After initially focusing on developing its own LPUs, Groq is now pivoting to provide cloud services using Nvidia's infrastructure, positioning itself within Nvidia's expansive AI ecosystem.
- Inference demand surge - As enterprises increasingly scale AI workloads, Groq aims to capitalize on the high demand for AI inference, although the long-term profitability of neoclouds remains uncertain.
The Bigger Picture: Groq's strategic pivot reflects a broader trend in the AI infrastructure landscape, where companies are increasingly reliant on established players like Nvidia for hardware while attempting to carve out niches in cloud services. This shift may lead to a consolidation of power within the AI infrastructure market, favoring those who can effectively leverage existing technologies rather than developing new ones. However, the long-term viability of neoclouds remains in question, as high capital expenditures and reliance on rapidly depreciating hardware could challenge profitability, potentially leading to a shakeout in the sector over the next year.
What is an EORβand why are companies using it?
Opening entities in every country can be slow, expensive, and hard to scale.
That's why more companies are using EOR to hire globally faster.
See how Oyster helps teams hire, pay, and support talent in 180+ countries while staying compliant along the way.
ποΈ AI Bytes
π€ Google announces Gemini 3.7 Flash just three weeks after previous release
Google has launched the Gemini 3.7 Flash model, just three weeks after the previous 3.6 Flash release, featuring core optimizations that enhance coding and agentic performance. With significant improvements in various benchmarks, including a notable increase in coding accuracy, Google aims to attract developers with a lower introductory price of $0.75 per million input tokens. However, the rapid release raises questions about the necessity of a new model so soon, especially as Google faces competition from rivals like OpenAI and Anthropic.
β‘ NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart
NVIDIA has launched Nemotron 3.5 Lightning on Amazon SageMaker JumpStart, a high-performance model optimized for high-volume agentic workloads, offering up to 4x higher throughput and 30% faster task completion. This open model, featuring a hybrid Mixture-of-Experts architecture, allows for easy deployment without manual infrastructure setup, making it suitable for various enterprise applications such as personal assistants and financial services. Organizations can customize and post-train the model to fit their specific needs while maintaining control over the resulting weights.
π€ Gemini becomes Google's fastest-growing product ever as it hits 1B users
Google's Gemini has achieved a remarkable milestone, reaching 1 billion monthly active users, making it the fastest-growing product in the company's history. Integrated across various Google services, Gemini's popularity is bolstered by its accessibility on both Android and iOS platforms, though concerns about its performance and internal leadership changes may pose challenges ahead. As the AI landscape evolves, Google must ensure Gemini remains competitive to sustain its user base and address emerging issues.
π’ Nvidia discloses $21bn stake in SpaceX
Nvidia has revealed a substantial $21 billion stake in SpaceX, marking a significant investment in the aerospace sector. This move underscores Nvidia's strategic interest in expanding its influence beyond traditional tech boundaries and into the realm of space technology, potentially leveraging AI advancements for aerospace applications.
π οΈ Top AI Tools This Week
Ollama
Bring open models into your workflow. Launch agents from the Ollama CLI, or connect editors and frameworks through Ollama's API.
Langfuse
Trace, evaluate, and improve AI agents with one open platform. Use production data to understand behavior, collaborate on fixes, and ship better quality at lower cost and latency.

