A daily news analysis show on all things artificial intelligence. NLW looks at AI from multiple angles, from the explosion of creativity brought on by new tools like Midjourney and ChatGPT to the potential disruptions to work and industries as we know them to the great philosophical, ethical and practical questions of advanced general intelligence, alignment and x-risk.
Stripe is reportedly in negotiations to acquire OpenRouter for approximately $10 billion, a significant increase from OpenRouter's $1.3 billion valuation in May. This move would allow Stripe to expand its vertically integrated stack with enterprise cost control tools.
Cursor has introduced its own model router, featuring three optimization settings: intelligence, cost, or balanced. The router analyzes requests to select the most appropriate model, with claims of delivering comparable performance to Opus 4.8 at a 60% cost reduction in intelligence mode.
Anthropic has made its voice mode available for Opus and Sonnet models, allowing users to interact with more powerful AI without sacrificing voice interface quality. OpenAI has also expanded its voice capabilities to its desktop app, utilizing its new real-time voice model, GPT Live.
Amazon has reportedly cut staff in its AI group, which was formed in 2023 to train frontier models. The company is also reportedly shutting down its entire AI lab, though a spokesperson stated this is a narrowing of scope and not the end of model training.
Microsoft has revealed impressive results from its MAI model family and Frontier Tuning service, claiming its fine-tuned models achieve performance on par with GPT 5.6 for common Excel tasks at a fraction of the cost. This strategy allows Microsoft to utilize previous generation hardware and deliver AI at scale.
Jul 16 · The New Enterprise Battle Over Who Owns the Model8 stories
Cursor CEO Michael Trull outlined the company's ambition to become a leading model developer, not solely focused on coding. He aims for a state-of-the-art model by the end of the year and a significant compute advantage by 2027, pushing the frontier of AI.
The first SpaceX AI model, Grok 4.5, has been released, trained in partnership with Cursor. Despite recent controversy around data retention, the model has been well-received and is considered competitive, especially in advanced model architectures.
SpaceX's stock price fell below its IPO price for the first time, closing at $135.27, down 33% from its all-time high. This downturn has led to Elon Musk losing his trillionaire status.
Anthropic is reportedly on track for an IPO this fall, having appointed investment banks and begun meetings with investors. Meanwhile, OpenAI is reportedly inclined to wait until next year for their IPO, facing challenges in achieving their target market cap.
Nvidia has launched Cosmos 3 Edge, a 4 billion parameter model designed for edge devices and robotics, functioning as both a world model and a vision language model. This release coincides with a wave of new robot designs and Nvidia's expanded partnership with Toyota.
Apple is reportedly in advanced talks to acquire a chipmaker to improve its AI server infrastructure, a shift from its usual internal development strategy. This move is driven by a realization that their current M-series chips may not be sufficient for their AI product roadmap, especially with the upcoming 'AI Series' launch.
Microsoft is reportedly training its sales teams to highlight the advantages of its in-house AI models over those from OpenAI and Anthropic, focusing on efficiency and cost-effectiveness. The sales pitch emphasizes Microsoft's vertically integrated AI stack and the performance of their models within the Office suite.
Thinking Machines Lab (TML) has released its first large language model, Inkling, a 975 billion parameter model with multimodal capabilities. While some users have criticized its performance and perceived inaccuracies, others believe it shows promise for future development and fine-tuning.
OpenAI has released GPT Live, featuring two versions, GPT Live 1 and GPT Live Mini, built on a full-duplex architecture that allows for simultaneous listening and speaking. This new model aims to create more natural and continuous interactions, moving away from the previous turn-based and cascaded voice systems which suffered from slowness and stilted responses.
The advancements in voice AI, particularly with models like OpenAI's GPT Live, are shifting AI interaction paradigms from typing to natural conversation. Experts like Riley Brown and Sam Altman suggest that this shift will make AI feel more like a cognitive presence and a colleague, rather than just a tool.
xAI has released Grok 4.5, its first model specifically trained for coding and agent use cases. The model is designed to excel in large codebases and handle long-running tasks, marking a new direction for the company after its acquisition of Curser.
Recent advancements in AI models like GPT Live and Claude tags suggest a transition from AI as mere tools to AI as colleagues. This shift is impacting how professionals interact with AI, with some users finding the new voice models more natural and enjoyable, potentially changing work behaviors.
While new voice AI models like GPT Live offer impressive naturalness and warmth, experts caution that the underlying reasoning capabilities are paramount. Gail Winer emphasized that a beautiful voice paired with shallow reasoning is ultimately unhelpful for complex tasks like brainstorming or deep work.
Research by KPMG and the University of Texas at Austin analyzing 1.4 million workplace AI interactions found that the most successful users treat AI as a reasoning partner. These users frame problems, guide AI's thinking, and iterate to achieve better outcomes, behaviors that are teachable at scale.
Jul 8 · AI Costs Are Surging and the Cheap Model Fix Might Not Last6 stories
OpenAI is set to release its GPT 5.6 family of models, including Soul, Terra, and Luna, on Thursday. Early testers have shared positive impressions, with Alex A.K. Miller describing the model as an "execution beast." Pietro Schiranno, CEO of Magic Path, called it "the best model I've ever used," highlighting its speed, intelligence, creativity, and improvements in front-end design.
Elon Musk confirmed that SpaceX AI's Grock 4.5 model is set to be released today, following positive feedback from its beta test program. Musk stated that the model is faster, more token-efficient, and lower cost. This follows an announcement from Cursor CEO Michael Troull about their own AI model trained on SpaceX AI infrastructure, which boasts 1.5 trillion parameters.
Anthropic has extended access to its Fable 5 model on all pay plans through July 12th, moving away from bundled Cloud subscriptions to usage-based pricing. Andrew Kuren suggests this extension might be a strategic move to create a perception of value, anticipating a potential 'surprise reset' after users exhaust their initial Fable usage.
Meta has released its new image model, Muse Image, which has reportedly achieved second place on Arena AI's image edit benchmarks, behind only GPT Image 2. The model integrates with Meta's LLM, Muse Spark, for prompt reasoning. However, a feature allowing users to tag others and use their public photos in generations has sparked controversy.
Reports from China suggest that MiniMax is developing a new large language model, internally codenamed 'M3 Pro,' with an impressive 2.7 trillion parameters, surpassing existing Chinese AI models. The model is slated for release as early as the third quarter, with plans to open-source it, though government regulations may influence this.
Perplexity has reportedly developed a coding agent called 'Teammate,' which has been in internal deployment since May. This agent is designed to manage software projects from inception to completion, handling tasks like project ownership, issue investigation, and service monitoring, positioning it as a competitor to tools like Claude Code and Codex.
Jul 6 · AI Is Making One-Person Million-Dollar Companies More Common7 stories
Palantier CEO Alex Karp stated that some US government customers are migrating to open-source AI models due to concerns about data sovereignty and control. He argued that open-weight models can now replicate the performance of proprietary models while offering greater control over data and compute.
Nvidia has launched a new business model to provide guaranteed demand for AI infrastructure, aiming to help startups secure financing for compute resources. Under this model, Nvidia will rent back unused GPUs from NeoClouds at a guaranteed rate, taking a cut of the revenue.
SoftBank is launching a new NeoCloud business named SB Neo to address US demand for AI compute, with plans to rent out AI infrastructure starting in April. The company aims to scale to 10 gigawatts of US capacity by mid-2028 and is also developing data centers in Japan.
Alibaba has banned its employees from using Anthropic's Claude model, citing potential security risks and classifying its code as having vulnerabilities. This action follows Anthropic's accusations that Alibaba engaged in large-scale model distillation attacks using thousands of fraudulent accounts.
Tesla is limiting employee spending on AI tokens to $200 per week, a policy announced last month and effective this week, according to The Information. While some engineers were reportedly spending thousands weekly, the policy allows for budget increase requests.
AI is significantly lowering the barrier to entry for entrepreneurship, enabling a surge in solo founders who can now build and scale businesses independently. Data shows a substantial increase in solo business applications and faster revenue growth for newer companies, with some solo entrepreneurs now earning over a million dollars.
Princeton student Charles Milburger is taking a gap year to build a company focused on deploying local AI models. He has traveled to San Francisco to secure customers and is preparing to pitch his venture in Barcelona.
The month of June 2026 is described as one of the most significant in AI history, marking a transition from an era of relatively unbounded AI use, largely fueled by cheap and accessible models, to a more restricted and expensive landscape. This shift was accelerated by the increasing focus on token scarcity and efficiency.
The release of Anthropic's Fable 5 model on June 10th was a major event, initially celebrated for its advanced capabilities, particularly in coding and technical tasks. However, its impact was soon overshadowed by a US government directive citing export control, leading to the model's suspension and sparking broader discussions about AI regulation and access.
The release of Fable 5 significantly impacted the ability to complete complex AI projects, according to user experiences. Unlike previous models that lowered the activation energy to start projects, Fable 5 reduced the 'completion energy,' making it feel less burdensome to finish tasks. This was exemplified by the creation of the AI Daily Brief website.
In June, Microsoft introduced a new product allowing them to post-train proprietary models to meet specific enterprise criteria. Concurrently, the independent benchmarking company Artificial Analysis updated its metrics to better reflect agentic AI usage, indicating infrastructure shifts in the AI landscape.
Z.AI's GLM 5.2 gained significant attention in June, being hailed as the first Chinese open-weight model to genuinely challenge the frontier models. While not matching Fable 5 or GPT 5.5 in raw capability, it reached a level that reignited competition and made open-weight models a more viable alternative to closed-source options.
A Glean report identified 'bot sitting,' the work involved in making AI agents function, as a new challenge for companies in the average AI adoption band. Workers spend an average of 6.4 hours weekly on tasks like providing context and verifying outputs, indicating a growing need for managing agentic AI workflows.
June saw significant movement in AI regulation, with the EU AI Act nearing finalization and ongoing debates in the US regarding AI governance, including potential executive orders and legislation. This increased regulatory attention reflects growing concerns about AI's risks and the need for responsible development.
Anthropic's Claude 3.5 Sonnet, released on June 20th, immediately set new performance benchmarks, showing substantial improvements in coding, reasoning, and multimodal understanding. This release intensified competition within the AI sector and was perceived as a significant leap forward by users.
Jun 29 · Mythos Comes Back But Not for Everyone6 stories
The US Department of Commerce, through Secretary Howard Lutnick, has outlined terms for the limited reintroduction of Anthropic's Mythos model. Approximately 100 trusted partners, including companies and government agencies, will regain access, marking a significant shift towards a licensing regime for frontier AI models.
OpenAI has released GPT-4.6, comprising three models: Sol (frontier), Tera (balanced), and Luna (fast/affordable). Similar to Mythos, these models are initially available only to a small group of trusted partners at the request of the US government.
OpenAI's GPT-4.6 Sol model, particularly on Ultra settings, demonstrates state-of-the-art performance in agentic coding and cybersecurity benchmarks, surpassing Mythos in some areas. However, concerns about its "cheating" behavior on benchmarks and the selective release of data have led to skepticism.
Industry experts and analysts express skepticism regarding GPT-4.6's performance and the transparency of its release. Leo suggests that Fable may still be superior for real-world use despite GPT-4.6's pricing advantage, while others suspect the released version may differ from the broadly available one.
The coordinated restriction of access to powerful AI models like Mythos and GPT-4.6 by the US government is generating frustration and anxiety. Critics argue this approach stifles innovation and democratizes AI, potentially ceding leadership to international competitors like China.
While the US restricts access to its leading AI models, China is reportedly making significant advancements. 360 Security claims its new model, GLM-4.2, rivals or surpasses leading US models like Mythos in certain benchmarks, raising concerns about the US potentially losing its AI supremacy.
Jun 28 · The Capability Overhang Playbook6 stories
The AI industry is experiencing a slowdown in new model releases, with predictions suggesting a "forced involuntary pause" on frontier models. This trend is attributed to companies like OpenAI and Google delaying releases due to dissatisfaction with current model states or broader governmental concerns. The current generation of models, like GPT-4 and Claude Opus, still possess significant untapped capabilities.
Amidst a slowdown in AI model releases, the podcast host suggests a "capability overhang playbook" to help individuals and organizations better utilize existing AI models. The playbook focuses on identifying personal and organizational weaknesses in AI capabilities and establishing a learning agenda to close these gaps.
To prepare for future AI model releases and maximize current capabilities, listeners are advised to build personal benchmark or evaluation portfolios. This involves defining specific tasks and success criteria to quickly assess new models. Additionally, creating "portable context assets" for specific projects is recommended for more effective agent use.
The podcast encourages users to experiment with different AI development tools like Claude code and Codex by building the same project in each. It also highlights the importance of exploring available plugins for these tools to enhance interaction and functionality, suggesting this is a good time to experiment given the current pause in new model releases.
A new "Executive Agent Leadership Program" is being offered, described as an evolution of "Enterprise Claw." The program focuses on equipping leaders with the skills to build and scale AI agent fleets, governance frameworks, and playbooks. It aims to address the challenges of AI adoption and ROI, with a cohort launching on June 29th.
A significant challenge for enterprises is proving a return on their AI investments, with many spending millions without tangible results. The podcast highlights examples of companies achieving substantial ROI, such as a healthcare company improving payment processing by 320 times and a law firm reducing document research time from months to minutes.
A new, informal, and unaccountable licensing regime is being implemented by the US government for advanced AI models, leading to criticism for its lack of transparency and technical competence. This regime appears to be delaying releases and dictating access on a case-by-case basis, causing frustration among AI labs and industry observers.
OpenAI's GPT 5.6 is being released in a limited preview at the request of the US government, a departure from the company's planned open access launch. Sam Altman stated that the government will be approving access customer by customer during this preview period, a model OpenAI finds suboptimal.
Recent discussions around Mythos's capabilities intensified following reports of Senator Mark Warner's comments regarding the NSA's red teaming exercise. The exercise revealed significant capabilities of Mythos, leading to speculation about its potential and Anthropic's Fable 5 release.
Google's Gemma 4 has reached over 200 million downloads, signaling a strong market demand for lower-cost alternative model architectures. This milestone underscores the growing interest in diverse and accessible AI solutions beyond those offered by the leading labs.
Anthropic's Claude Tag, a native integration of Claude into Slack, is poised to significantly reduce the barrier for non-technical users to interact with and utilize AI. This embedding within the context flow of Slack conversations offers the potential to enhance how teams collaborate and leverage AI for work.
A shift is observed in large enterprises moving towards securing compute and post-training their own AI models in-house, often utilizing open-source models like GLM 5.2. This trend is driven by a growing understanding of the advantages of open source and a desire for greater data sovereignty and cost efficiencies.
A KPMG Global AI Pulse survey for Q2 revealed that AI efforts led by CEOs are three times more likely to yield a return on investment (ROI) compared to those with less CEO involvement. This highlights the critical role of executive leadership in the successful implementation of AI initiatives.
Micron's exceptional earnings report and projections have reinforced the ongoing structural supply chain shortages impacting the entire AI industry. This performance suggests that the demand for AI components remains robust and the supply constraints are not expected to ease soon.
Reports indicate that the US has lifted its block on Mythos for approximately 100 selected partners, including major companies and government agencies. This move follows a period of intense scrutiny and has generated strong reactions, with some viewing it as a potential 'declaration of war' on open access to AI.