TLDRocket
Sign in
Latest Open-weight AI models are catching up to the frontier. The safety gap... — TechCrunch AI Anthropic signs $10B deal with AI cloud startup Volta — TechCrunch AI Meet Wrinkles, an app that uncovers the hidden stories of the places a... — TechCrunch AI Nvidia doesn’t mess around: A week after open AI industry group formed... — TechCrunch AI PipeNetwork/minimax-h3-mlx — Simon Willison Introducing Web Search on Amazon Bedrock for foundation model groundin... — AWS Machine Learning Tino Cuellar joins Anthropic as Chief Global Affairs Officer — Anthropic News Today’s Codex will feel “primitive” by fall — and its own team’s roadm... — The New Stack

Every AI story that matters — and the intelligence behind it.

TLDRocket reads 60+ sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

AI Market Index

51 ▲ 4

: 3 : 6 : 5 : 5 : 5 : 11 : 18 : 24 : 70 : 49 : 47 : 51

#1 AI Momentum

OpenAI

Weekly ranking →

Latest funding

$150M

HappyRobot →

Tracked now

5,068

Profiles · 419 events →

View:

Today

30-second scan All events →
  1. 1
  2. 2
  3. 3
  4. 4
  5. 5

Nvidia doesn’t mess around: A week after open AI industry group formed, it’s already showing progress

TechCrunch AI 1 hour ago 62 sources

Nvidia-led Open Secure AI Alliance, formed a week ago with over 120 member companies, has already launched a working group called SAFE to develop AI cybersecurity guidelines including confidential incident reporting and blame-free analysis protocols. Members are contributing open source tools such as Nvidia's Garak vulnerability scanner, Amazon's Strands Agents, and Okta's agent identity technology to create a shared security framework. The group aims to build an open source ecosystem for securing AI agents in enterprises, though notable absences include Anthropic, OpenAI, and Google despite their support for the original letter.

Trending stories

Business & funding

Beyond the headlines

Every story feeds a living map of the AI industry.

Briefings for your role: CEO CFO COO CTO CISO CMO

Also tracked: Regulations Industries Physical AI Conferences

Tuesday, 4 August 2026

Open-weight AI models are catching up to the frontier. The safety gap remains.

TechCrunch AI 1 hour ago 15

China's GLM-5.2 open-weight AI model has narrowed the capabilities gap with frontier models like GPT-5.5 and Claude Opus 4.7 in cyber and biological tasks, but lacks comparable safety measures. GLM-5.2 refused none of the offensive tasks SaferAI tested, while Claude Opus 4.7 consistently refused such requests. The divergence between capability and safety in open-weight models creates a governance challenge as these powerful systems become more accessible without enforceable safeguards.

Anthropic signs $10B deal with AI cloud startup Volta

TechCrunch AI 1 hour ago 41

Anthropic signed a $10 billion compute deal with Volta, a newly founded cloud startup, to receive cloud capacity over six years from a Norway data center built with Bitdeer and powered by Nvidia's Vera Rubin chips at 133 megawatts. Anthropic secured this agreement as part of an aggressive expansion of compute resources to compete with rivals, following similar recent deals with SpaceX and Amazon. The arrangement gives Volta a major client and establishes Bitdeer's involvement in developing AI infrastructure outside crypto mining.

Meet Wrinkles, an app that uncovers the hidden stories of the places around you

TechCrunch AI 1 hour ago 36

Wrinkles is a location-based app that uses AI to automatically surface stories and historical information about places around users as they explore. The app has built a global foundation of 1.3 million points of interest across 177 countries, with the ability for historians, museums, creators, and brands to add their own content. The company monetizes through partnerships with museums, tourism boards, and businesses rather than charging users, planning to become a standard tool for discovering local history during commutes, travel, and everyday exploration.

Nvidia doesn’t mess around: A week after open AI industry group formed, it’s already showing progress

TechCrunch AI 1 hour ago 28 62 sources

Nvidia-led Open Secure AI Alliance, formed a week ago with over 120 member companies, has already launched a working group called SAFE to develop AI cybersecurity guidelines including confidential incident reporting and blame-free analysis protocols. Members are contributing open source tools such as Nvidia's Garak vulnerability scanner, Amazon's Strands Agents, and Okta's agent identity technology to create a shared security framework. The group aims to build an open source ecosystem for securing AI agents in enterprises, though notable absences include Anthropic, OpenAI, and Google despite their support for the original letter.

PipeNetwork/minimax-h3-mlx

Simon Willison 2 hours ago 1

PipeNetwork ported MiniMax's new H3 multimodal model to MLX for Apple Silicon, enabling text-to-video generation on consumer Macs. The model generated a 15-second video from a text prompt in 45 minutes on an M5 Max MacBook Pro after downloading 115 GB of weights. Users can now run video generation locally on Apple hardware, though output quality depends on detailed prompt specifications including audio guidance.

Introducing Web Search on Amazon Bedrock for foundation model grounding

AWS Machine Learning 2 hours ago 12 2 sources

AWS announced general availability of Web Search on Amazon Bedrock, a server-side tool that grounds foundation model responses in current web knowledge without requiring third-party vendor integration. The feature is available through the OpenAI Responses API with a single parameter, supporting semantic snippet extraction from Amazon's web index and knowledge graph while keeping data within AWS infrastructure. Developers can now add web grounding to AI applications with minimal setup, reducing hallucinations and enabling models to answer questions about recent events beyond their training data.

Tino Cuellar joins Anthropic as Chief Global Affairs Officer

Anthropic News 16

Mariano-Florentino Cuéllar, former president of the Carnegie Endowment for International Peace and former California Supreme Court Justice, has joined Anthropic as its first Chief Global Affairs Officer to lead policy, government relations, and international engagement work. Cuéllar brings decades of experience across law, technology, international security, and public institutions, including prior roles directing Stanford's international security centers and serving on presidential intelligence and state department boards. His appointment positions Anthropic to deepen relationships with governments and policymakers as they develop frameworks for governing artificial intelligence globally.

Today’s Codex will feel “primitive” by fall — and its own team’s roadmap backs it up

The New Stack 2 hours ago 33

OpenAI's Thibault Sottiaux predicted that Codex will become outdated within 2-3 months as the company moves toward more advanced AI agents requiring persistent cloud infrastructure rather than local machines. OpenAI's planned acquisition of Ona, which provides secure cloud development environments used by 2 million developers, would enable agents to work continuously in customer clouds without depending on a developer's laptop being online. This shift requires new infrastructure management layers including agent identities, access controls, and logging systems to keep pace with autonomously operating coding agents.

OpenAI’s Astra just proved 10 long-standing math and science theorems. The tokens cost $2,000.

The New Stack 3 hours ago 32 5 sources

OpenAI's internal Astra model generated machine-verified proofs for 10 long-standing mathematical and theoretical computer science problems. The company estimated the inference cost at approximately $2,000 in GPT-5.6 Sol API tokens, marking the first time it quantified the cost of frontier reasoning work. This pricing framework could enable smaller research teams to access frontier AI reasoning through APIs rather than building their own models, shifting from training costs to inference budgets as the relevant planning metric.

SpaceX has more neocloud revenue

The Verge 4 hours ago 30

SpaceX's AI revenue tripled to $2.6 billion, driven by compute deals with Anthropic and Google that position the company against competitors like CoreWeave. The AI division lost $1.5 billion this quarter, down from larger losses in the prior year. SpaceX is shifting toward providing infrastructure services to other AI companies rather than developing AI products itself.

Microsoft Tells Engineers ‘Tokenmaxxing Is Not What We Are Optimizing For’

404 Media 5 hours ago 1

Microsoft has introduced spending limits on AI tools for employees and designated GPT-5.6 as the default model, directing engineers to focus on productivity gains rather than maximizing AI usage. Starting July 2026, Microsoft divisions will have AI token budgets, with current employee spending ranging from hundreds to thousands of dollars monthly. The policy reflects a broader industry trend of controlling AI costs as companies recognize that higher token consumption does not always deliver proportional business value.

Automated web insight extraction with Amazon Bedrock AgentCore

AWS Machine Learning 5 hours ago 34 2 sources

Amazon published a tutorial on building an automated web insight extraction system using Bedrock AgentCore Browser, which renders JavaScript-heavy pages reliably and extracts insights with AI models. The system monitors RSS feeds, retrieves full page content using managed browser sessions (taking 10–30 seconds per render), and indexes results in OpenSearch Serverless for semantic search. Teams can now monitor competitors, industry trends, and regulatory changes without manual website checks, with safeguards applied through Bedrock Guardrails to filter harmful content before indexing.

NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US

NVIDIA 5 hours ago 15

NVIDIA is participating in the NSF's State and Regional AI Infrastructure Hubs program, which launched today to expand AI computing access and education across US universities and regional consortia. The University of Florida model from 2020, which grew to over 300 AI-focused faculty and $511 million in research awards, will serve as a national template for the program. Regional hubs will enable institutions to share computing resources, develop workforce training pathways, and connect AI research infrastructure to local economic priorities and employer needs.

Claude, Gemini, and GPT-5 can handle every SDLC task. Almost none of them should.

The New Stack 5 hours ago 35

Large language models like Claude and GPT-5 can handle every software development lifecycle task, but organizations should use specialized smaller models instead for most work to control costs and improve governance. A single frontier model requires more resources and scales less efficiently than a specialized AI supply chain where different models handle distinct stages like code generation, testing, and compliance. As AI moves deeper into production software delivery, governance and orchestration layers must be built into the pipeline from the start rather than added afterward to track usage, costs, and policy compliance.

Spotify expands AI remix and covers project with Merlin partnership

TechCrunch AI 5 hours ago 11

Spotify announced a partnership with Merlin to expand its AI-powered covers and remixes product, which will allow fans to create derivative works from participating artists' music with consent and compensation. The deal adds over 30,000 independent labels to the program, and Spotify plans to launch a research preview to a subset of users as a paid add-on. The move positions Spotify as offering the only legal way for artists to participate in AI music generation for interactive platforms.

Texas halts new data centers as governor calls for audits

TechCrunch AI 5 hours ago 25 2 sources

Texas Governor Greg Abbott announced that all new data center projects must undergo audits by state regulators, citing concerns about grid capacity and rising electricity prices. ERCOT's interconnection queue has grown to 474 gigawatts of pending projects—more than 90% data centers—more than doubling from 233 gigawatts in January, representing over five times the grid's peak demand. The audits could slow or halt Texas's role as a major data center hub if regulators determine projects pose risks to grid reliability or electricity costs.

Elon Musk spends half his time talking robots and AI on Tesla earnings calls

TechCrunch AI 5 hours ago 30

Elon Musk now spends nearly 50% of his remarks on Tesla earnings calls discussing artificial intelligence, robotaxis, and the Optimus robot, up from 15-20% in 2022. Analysis of seven years of earnings call transcripts shows Musk's focus on AI and robotics has intensified as Tesla's car sales growth stalled, with robotics mentions rising from 2% in 2022 to 10% recently and reaching nearly a third of his remarks in Q3 2025. Other Tesla executives remain more focused on the automotive business at roughly 30% of their discussion time, indicating a divergence between Musk's AI-focused narrative and the company's actual revenue sources.

NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use

NVIDIA 6 hours ago 35

NVIDIA released Alpamayo 2 Super, an open-source reasoning model for autonomous vehicles, now available for commercial use under a permissive Linux Foundation license. The model ranks first on LingoQA autonomous driving benchmark, outperforming Gemini 2.5 Pro by 15.1 points and GPT-4o by 23.2 points, with 3x the scale of earlier Alpamayo versions. Developers can now deploy the model commercially without additional licensing while building specialized AV systems with better reasoning transparency and safety validation aligned to ISO standards.

As AI Increases Demands on Memory, Storage Steps Up

NVIDIA 6 hours ago 33

NVIDIA is unveiling storage infrastructure advancements designed to handle surging AI data demands, including the Vera CPU which delivers 3.21x higher throughput than x86 processors in compression and encryption pipelines. The company is open-sourcing cuFile APIs to enable GPUs to read and write directly to storage, and launching the Storage-Next initiative with over 40 vendors to align on GPU-driven storage standards. These changes shift storage from a passive repository to an active part of the data path, allowing AI systems to access data in microseconds rather than minutes and reducing the gap between AI computing capacity and available memory.

Why R&D Waste Persists Despite Widespread AI Adoption

IEEE Spectrum AI 6 hours ago 23

A whitepaper examines how R&D waste persists even as organizations adopt AI, finding that over one-third of companies waste 25–40 percent of R&D budgets on projects never reaching market. The median cost of a failed project during development is over one million dollars. Organizations apply AI mainly to execution tasks like data analysis rather than decision support, leaving the early ideation phase—where intelligence would prevent waste—unaddressed.

What developers really think about Qwen3.8-Max: “An API business model wearing an open source jacket”

The New Stack 6 hours ago 9 6 sources

Alibaba launched Qwen3.8-Max, a 2.4 trillion parameter multimodal model with a 1 million token context window, and promised to open-source its weights. Developers questioned whether promised weights had actually shipped alongside the launch and skeptically evaluated self-benchmarked performance claims versus third-party validation. The real value hinges on robust operational infrastructure, cost efficiency, and whether enterprises gain meaningful control over long-horizon agent deployments without constantly rebuilding context or introducing silent failures.

Protected: Assessing Sovereign AI: A Two-Pronged Framework

CSET Georgetown 6 hours ago 47

This is a metadata page from the Center for Strategic and Emerging Technology (CSET) describing their research mission on technology and security policy, not an article with substantive content. CSET publishes reports and policy analysis on emerging technologies including AI and semiconductors. The page mentions a Wall Street Journal article by staff member Jacob Feldgoise about U.S. semiconductor manufacturing support, but does not present the actual research or findings.

Apple says more ex-employees may have taken confidential data to OpenAI

TechCrunch AI 7 hours ago 44 4 sources

Apple is seeking a preliminary injunction against OpenAI in a trade secrets case and has identified 11 additional former Apple employees who may have been involved in transferring confidential data beyond the two originally named defendants. The filing reveals specific incidents including former employees sharing Apple proprietary information about unannounced products and taking screenshots of confidential documents before interviewing at OpenAI. OpenAI denies possessing any Apple trade secrets and has disputed Apple's claims, pointing to procedural errors Apple made during its investigation.

Nvidia’s NOOA makes an agent one Python class

The New Stack 7 hours ago 44

Nvidia released NOOA, a framework that represents AI agents as single Python classes, consolidating capabilities, state, and prompts into one unified structure to reduce fragmentation in agent development. On SWE-bench Verified, NOOA achieved 82.2% accuracy with GPT-4o using 29 LLM calls and roughly 1.1M tokens per task, compared to competitors requiring 66 calls and 2.2M tokens for 78.2% accuracy. The centralized approach makes agents easier to inspect and audit but concentrates security risks and may blur distinctions between deterministic and probabilistic code paths.

Deploy local agents everywhere with LFM2.5-2.6B

Hugging Face Blog 7 hours ago 4

Liquid AI released LFM2.5-2.6B, a 2.6-billion-parameter language model designed to run AI agents locally on consumer devices like laptops and phones while supporting tool calling and multi-step workflows. The model achieves 220 tokens per second on an Apple M5 Max and 113 tokens per second on an AMD Ryzen CPU while using under 2.5 GB of memory, and performs competitively with models 4 times larger on instruction following and tool-use benchmarks. Developers can now deploy capable agents entirely on-device without cloud inference costs, keeping user data private while maintaining agentic reasoning capabilities.

OpenAI says Apple's trade secrets lawsuit is "aggressive and oddly personal"

Ars Technica 7 hours ago 16 4 sources

OpenAI published a blog post denying Apple's lawsuit allegations that it stole trade secrets related to hardware designs for AI devices, calling the suit careless and personal. Apple filed the lawsuit last month claiming OpenAI misappropriated confidential information as OpenAI prepares to launch consumer AI hardware. The dispute marks an ongoing escalation in competitive tensions between the two companies over AI product development.

‘Not healthy’ LLM use is more common than you think

The Verge 7 hours ago 40 2 sources

Hank Green paused his YouTube production after criticism over using large language models for research sourcing, acknowledging the practice as unhealthy. Green clarified he used AI to find sources rather than write scripts, but the controversy highlights tensions between creator authenticity and AI tools trained on uncompensated work. The incident raises questions about where creators should draw lines with AI use without compromising their credibility.

A16z backs HappyRobot at $1.2bn valuation

Sifted 7 hours ago 40 2 sources

HappyRobot, an enterprise AI agent startup, raised $150 million in Series C funding at a $1.2 billion valuation led by Prysm Capital and Eurazeo. The company, founded in 2022 in Madrid, works with over 150 businesses including DHL and Uber to build and deploy AI agents across complex workflows. The funding will support expansion of its AI platform capabilities and hiring across engineering, deployment, and sales teams.

What my agent knows about me

Ben's Bites 8 hours ago 23 2 sources

The author tested a 'reflection engine' prompt with AI agents and found that Sol Max produced a more coherent analysis of their personal data and memories than Fable High. OpenAI cut GPT-5.6 Luna's price by 80% and introduced the Astra model, which solved 10 math and theoretical computer science problems. Several new AI models launched with lower pricing and new capabilities, enabling users to accomplish significantly more work at reduced costs.

AI Leaders Propose SAFE Guidelines for Cybersecurity Transparency

NVIDIA 8 hours ago 15 62 sources

The Open Secure AI Alliance, comprising over 120 organizations, is developing SAFE (Shared AI Findings Exchange) guidelines to improve cybersecurity for agentic AI systems by enabling confidential incident sharing and collective defense. The framework includes contributions from NVIDIA, Cisco, CrowdStrike, Hugging Face, Red Hat, and newer members like Amazon and Visa, spanning identity controls, harnesses, runtime guardrails, specialized security models, and observability tools. Organizations can now share threat intelligence and security findings openly to accelerate ecosystem-wide protection against AI-specific attack surfaces.

AI agents can create database sprawl issues. YugabyteDB’s solution is more agents!

The New Stack 8 hours ago 19

YugabyteDB announced AMP, a serverless PostgreSQL tier designed to manage hundreds or thousands of databases created by AI agents, each with isolated data layers. The platform scales to zero when idle and charges by CPU minute when active, with four built-in AI agents (Architect, Voyager, Perf Advisor, Nexus) handling database creation, migration, performance monitoring, and ecosystem integration. Companies can now experiment cheaply with AI agents on small databases and scale to distributed PostgreSQL without application rewrites as workloads grow.

Astro’s GitHub issue backlog is heading to zero for the first time in 5 years. Now Cloudflare is open-sourcing the tool that did it.

The New Stack 8 hours ago 38

Astro, a JavaScript framework owned by Cloudflare, has reduced its GitHub issue backlog from over 200 to around 20 using AI agents, and expects to reach zero within a month—a first in the project's five-year history. The team built triagebot-action, a GitHub Action that runs a four-stage pipeline (reproduce, diagnose, verify, fix) with separate AI agents for each stage, powered by the Flue framework. Cloudflare open-sourced triagebot-action so other maintainers can automate issue triage, though adoption beyond Astro remains early and the tool's fix capability is still being refined.

Is the future of data centers portable? Runware builds a pod to find out

TechCrunch AI 8 hours ago 50

Runware, an AI infrastructure company, announced the Sonic Inference Pod, a modular transportable data center unit designed for AI inference workloads. The company currently operates 10 pods deployed across the U.S., Europe, and Asia-Pacific, with 160 available sites, following a $50 million Series A funding round in December. The pod design enables faster deployment, distributed computing closer to users, and avoids water cooling, positioning Runware as an alternative to traditional hyperscale data center expansion.

EU’s €5bn tech superfund officially open for business

Sifted 8 hours ago 41

The European Commission's €5bn Scaleup Europe Fund, run by EQT, has completed legal proceedings and begun investing after a year of preparation. The fund targets investments of €100m–500m per deal with a pipeline of over 100 startups, and is reportedly in talks to lead funding rounds for Mistral (€3bn target) and The Exploration Company (~$350m). European startups now have access to large-scale capital from a dedicated EU fund aimed at building world-leading companies within the continent.

HappyRobot lands $150M Series C to scale agentic AI for enterprise operations

Tech.eu 8 hours ago 1 2 sources

HappyRobot, an enterprise AI company building agents for supply chain and operational workflows, raised $150 million in Series C funding at a $1.2 billion valuation. The company serves 150+ customers including DHL and Uber, with one customer automating 28,000 work hours monthly and customer service agents achieving 9.4/10 satisfaction scores. The funding will expand HappyRobot's platform capabilities, enterprise integrations, and deployment teams across eight global offices.

Introducing Shieldstral.

Mistral AI 9 hours ago 50

A company released Shieldstral, a 3-billion-parameter open-source safety classifier for moderating text and images that accepts custom policies at inference time without retraining. The model matches the performance of guardrail systems up to 7 times larger and runs on a single 16GB GPU, using a question-answering approach where policies are supplied as plain-language prompts. Organizations can now adapt content moderation rules to different contexts and audiences dynamically instead of retraining fixed models for each deployment.

Exclusive: Nscale to provide compute to Oxford, UCL and Imperial AI researchers

Sifted 9 hours ago 16

Nscale, a London-based data centre builder, will provide $1 million worth of compute resources to AI researchers at Oxford, UCL, and Imperial College London. The agreement reflects growing demand for domestically controlled computing capacity among UK research institutions developing advanced AI systems. This arrangement allows British researchers to access sovereign compute infrastructure without relying on international cloud providers.

EON wants to move the data superhighway from ocean fiber to space lasers

TechCrunch AI 9 hours ago 43

Endeavor Optical Networks, a new startup, secured $10.75 million to build a satellite network using lasers to transmit data between data centers at speeds exceeding 200 terabits per second, replacing undersea fiber optic cables. The company plans to launch roughly 20 satellites with an initial demo spacecraft targeted for end of 2027, aiming for at least 800 gigabits per second optical downlink throughput. If successful, EON would offer hyperscalers and AI labs faster, more reliable intercontinental data routes, particularly on expensive or underserved corridors like France-to-Australia.

Texas says data centers must pass an audit before connecting to the grid

The Verge 9 hours ago 7 2 sources

Texas governor Greg Abbott directed state regulators to audit data centers before they connect to the grid, requiring disclosure of incentives received, grid dependency, water usage, and community impact plans. New data center proposals will face additional verification steps through the Public Utility Commission of Texas and ERCOT. The audit requirement could delay new facility approvals and may reduce the pace of data center expansion in the state.

Frame selection is the whole game: notes from making LLMs watch video

TLDR Dev 10 hours ago 24

A developer describes techniques for efficiently extracting keyframes from videos for LLM analysis, using adaptive scene detection and three-channel deduplication to select the most informative frames within a token budget of 100–150 frames per video. Key technical details include per-frame scene scores compared against rolling averages, RGB-based global deduplication at 16×16 resolution, a separate action channel that catches small subjects via 32×32 grid analysis, and a settled channel at 192×192 resolution for UI and text changes. The result is an open-source tool (claude-real-video) that outputs JPEGs, transcripts, and a manifest file, available as an MCP server, allowing models to analyze video content directly rather than relying on human summaries.

Next.js 16.3

TLDR Dev 10 hours ago 44

Next.js 16.3 shipped with Instant Navigations (a suite of SPA-like responsiveness tools), dev server memory cuts up to 90%, and 22% faster server-side rendering. The release includes faster builds via disk caching, TypeScript 7 integration, versioned documentation for AI agents, and experimental features like the Rust-based React Compiler. Developers get performance gains without code changes, plus new tooling for AI agents to access version-matched docs automatically.

Experiments with AI Code Review

TLDR Dev 10 hours ago 15

Wealthfront built an AI code review system called Iris that uses Claude Opus as the lead agent with adversarial sub-agents (prosecution and defense models) to evaluate potential bugs and issues in pull requests. The system achieved an average review cost of $4 and 10-minute turnaround, with engineer ratings skewing toward useful and great feedback while significantly reducing false positives compared to earlier attempts. The AI review complements rather than replaces human code review, running after self-review and before peer review to maintain reviewer independence.

What's the largest software project AI can complete on its own?

TLDR Dev 10 hours ago 30

Researchers introduced MirrorCode, a benchmark that tests AI models on reimplementing entire software programs end-to-end without access to original source code. Claude Opus 4.7 successfully reimplemented gotree, a 16,000-line bioinformatics toolkit in Go, in 14 hours at a cost of $251, a task that would take a human engineer 2–17 weeks. The benchmark's open-source release and leaderboard enable systematic measurement of AI capabilities on long-horizon coding tasks that require weeks of inference budget rather than dollars.

The Asus Chromebook Plus CX34 is at one of its lowest prices

The Verge 10 hours ago 6

The Asus Chromebook Plus CX34 laptop is available at discounted prices below $400 across multiple retailers. The most affordable option is $349.99 at Walmart with 128GB storage and a 13th Gen Intel Core i3 processor, while Best Buy offers a faster 13th Gen Core i5 model for $399.99. The price reductions make this three-year-old Chromebook a more accessible option for budget-conscious buyers.

The Sequence Knowlege #907: The Brain Transplant: Distilling Transformers Into Other Architectures

TheSequence 10 hours ago 16

Researchers are transferring knowledge from transformer models into fundamentally different architectures like state-space models and linear RNNs through cross-architecture distillation, a technique that preserves the capability of the original model despite changing its computational substrate. A key distinction is that previous distillation kept teacher and student in the same architectural family, but this approach breaks that assumption by using entirely different machine types. This capability transfer opens economic opportunities by allowing efficient non-transformer architectures to inherit transformer-level performance, potentially reducing computational costs in deployment.

What Are Companies Getting for All That AI Spending?

TLDR 10 hours ago 44

The Linux Foundation created the Tokenomics Foundation to establish shared standards for AI companies to report costs and benefits of their models, addressing the lack of consensus on energy requirements and model selection across thousands of different AI systems. The initiative aims to create common disclosure parameters so that costs and performance can be directly compared. This enables companies and customers to make more informed decisions about AI spending and model selection based on standardized metrics rather than proprietary claims.

The Endgame Of Vertical Integration

TLDR 10 hours ago 12

AI model companies like Anthropic are increasingly building their own application layers (harnesses) alongside their models rather than relying on third-party developers, forcing specialized agent labs to consider training their own models to remain competitive. The article cites Anthropic's inference margins improving from 38–40% in 2025 to over 70% currently as evidence that integrated model-harness development delivers superior performance and economics. Agent labs focused on specific domains like legal or finance must now choose between building custom models themselves or accepting diminished competitive positioning as foundation model companies dominate both the underlying intelligence and the application layer.

LLMs reward expertise

TLDR 10 hours ago 46

Large language models amplify the value of domain expertise rather than eliminating it, as demonstrated by mathematician Terence Tao's superior results with ChatGPT compared to non-specialists asking the same model. Tao's approach—short precise queries, identifying when outputs seem overcomplicated, pushing back without direct contradiction, and making independent suggestions—relies entirely on deep mathematical knowledge. Users with specialized knowledge in their field can steer LLMs toward better solutions by recognizing what good outputs look like and iterating strategically, whereas those without domain expertise can only accept whatever the model produces first.

Cloudflare Computer

TLDR 10 hours ago 43

Cloudflare introduced @cloudflare/computer, an open-source runtime package that gives AI agents their own virtual computer with a shared filesystem and multiple execution environments. The package optimizes agent scalability by running most tasks on Cloudflare's lightweight isolates (based on Workers technology) and spinning up containers on-demand for heavier compute needs, with the goal of keeping container usage below 10% of agent work. This approach addresses the industry's critical shortage of compute capacity for running millions of concurrent agents across all cloud providers.

Drug Discovery Has No Magic Wands

TLDR 10 hours ago 34

Daphne Koller argues that AI's promise to cure disease through superintelligence overlooks a critical bottleneck: we lack sufficient understanding of human biology to identify which disease mechanisms are worth targeting. While AI has excelled at molecular design (stage 2 of drug discovery), over 90% of drugs fail in clinical trials because researchers target the wrong biological mechanisms, not because molecules are poorly designed. Real progress requires massive new biological measurement efforts and human clinical data, which no amount of computational optimization can replace, making the current industry focus on better molecular design tools largely misguided.

Monava closes funding round as demand for passive drone detection grows

Tech.eu 11 hours ago 6

Monava, a Swedish-Finnish company specializing in AI-powered acoustic drone detection, has closed a funding round led by Gungnir Capital alongside Foundry Ventures and Hede Capital. Drones account for roughly 70 percent of all casualties in Ukraine's war, and Monava's passive detection system is already deployed across the Nordics and Ukraine. The funding will accelerate product commercialization and organizational scaling to meet rapidly growing demand from European defence customers.

The global grassroots gatherings trying to humanize the AI boom

Rest of World 11 hours ago 10

Grassroots organizations like AI Salon and AI Collective have launched chapters globally since 2023 to hold conversations about AI's societal impact beyond Silicon Valley's dominant narrative. AI Salon operates in three cities with ~20-person gatherings, while AI Collective has grown to 200,000 members across 200 chapters in 50 countries, mostly run by volunteers. These groups aim to broaden AI discourse and build local agency, though critics note participants still tend to be tech-industry workers reproducing existing power imbalances rather than reaching ordinary people.

How to build secure and user-friendly AI: ‘If data sits in a silo, it only sees part of the picture’

Sifted 12 hours ago 14

Box's VP of DACH discusses how companies can adopt AI securely by consolidating unstructured data and automating governance controls rather than choosing between security and usability. He emphasizes that scattered data limits AI effectiveness, citing examples like service technicians receiving outdated information when documents sit in separate systems. Organizations should prepare their data infrastructure and user permission systems before deploying AI models, reducing security risks and Shadow AI adoption where employees use unsanctioned tools.

Building an Advanced AI Skill Security Auditing Pipeline with NVIDIA SkillSpector, LangGraph, YARA Rules, SARIF, and CI Policy Gates

MarkTechPost 13 hours ago 17

NVIDIA SkillSpector is a security auditing tool that scans AI agent skills for vulnerabilities using LangGraph, YARA rules, and SARIF reports. The tutorial demonstrates scanning four synthetic skills—pdf-summarizer, repo-janitor, invoice-sync, and notes-mcp—with risk scores ranging from clean to malicious, including detection of embedded credential theft, command injection, and unapproved telemetry beacons. Results enable teams to suppress baseline findings, detect regressions, enforce CI security gates, and visualize fleet-wide risk distribution before deployment.

Legal AI startup Aavalynx raises £1.5M to cut the cost of corporate disputes

Tech.eu 13 hours ago 17

Aavalynx, a legal AI startup founded in 2023, raised £1.5 million in pre-seed funding to help enterprises analyze and manage corporate litigation disputes more effectively. The platform uses proprietary AI to structure dispute data, enabling companies to make earlier strategic decisions and reduce legal costs, with reported ROI of 30x in saved damages and legal fees. The funding allows the company to expand its product and team while positioning legal teams to shift from reactive dispute management to proactive risk assessment across their entire portfolio.

UK mulls making employers ask before installing bossware

The Register 13 hours ago 20

The UK government is consulting on whether to require employers to consult workers before deploying workplace monitoring technology including AI-powered productivity scoring and keystroke logging. One in three UK organizations currently monitor employees' digital activity, up from one in five two years earlier. If statutory rules are imposed, employers would face a new consultation requirement alongside existing GDPR and AI Act compliance, potentially slowing adoption of workforce monitoring tools.

OpenAI drags Apple’s lawsuit into the court of public opinion

The Verge 13 hours ago 27 4 sources

OpenAI published a blog post responding to Apple's trade secret theft lawsuit, characterizing it as careless and oddly personal while sharing iMessage and email exchanges to challenge Apple's allegations. OpenAI did not file a formal legal response but instead attempted to sway public opinion by highlighting contradictions in Apple's case. The public airing of communications and counter-narrative may influence how the dispute is perceived outside the courtroom.

Estel Technologies raises €270K to modernise tech staffing sales

Tech.eu 14 hours ago 30

Estel Technologies, a Bulgarian startup, raised €270,000 in pre-seed funding to develop an AI-powered platform that helps tech staffing firms identify and prioritize sales opportunities by combining demand verification, lead scoring, and buyer matching. The platform's Demand Confidence Score analyzes market signals and ranks opportunities for individual users, addressing inefficiencies caused by outdated job listings and fragmented workflows in the tech staffing industry. The funding will support core technology development, customer growth across European markets, and preparation for a seed round and US expansion in 2027.

Shiplog raises $1M to build AI customer intelligence for B2B SaaS

Tech.eu 14 hours ago 26

Shiplog, a Paris-based startup, raised $1 million in pre-seed funding to build Ada, an AI agent that delivers personalized customer experiences for B2B SaaS companies by continuously evaluating individual customers and recommending next-best actions across their lifecycle. The platform analyzed over four million events across 1,000 customers during pilot programs, demonstrating ability to operate at scale. Instead of treating customers as segments, Ada creates individualized profiles and personalizes marketing, product interfaces, and communications in real time, enabling companies to provide personalized engagement to thousands of accounts that typically wouldn't receive dedicated attention.

Exclusive: Plural backs $43m round for next-generation battery startup Ore Energy

Sifted 16 hours ago 13

Dutch startup Ore Energy raised $43 million in Series A funding to scale its iron-air battery technology, with investors viewing long-duration energy storage as critical for Europe's AI infrastructure needs. The round valued the company at $16 million post-money. The battery technology enables grid-scale energy storage to support renewable power systems serving data centers and other energy-intensive applications.

Y Combinator Open-Sources QM: An MIT-Licensed Multiplayer Agent Harness That Runs In Slack And The Web

MarkTechPost 17 hours ago 20

Y Combinator open-sourced QM, a multi-agent harness for workplace collaboration that runs in Slack and the web, under the MIT license. The system is designed for organizations of 10–500 people with at least one platform engineer, supports multiple AI models (Pi, OpenCode, Codex, Claude Code) without vendor lock-in, and isolates each user and room with separate memory, permissions, and sandboxed execution. Deployments can now use QM for internal tasks like inbox triage, document search, code testing, and project tracking without being tied to a single AI vendor.

[AINews] Qwen 3.8 Max(2.4T) and 27B, new open weights models for Coding and Cowork

Latent Space 17 hours ago 2 6 sources

Alibaba released Qwen 3.8 Max, a 2.4-trillion-parameter sparse model with open-weight versions promised, alongside a smaller 27B model, both available via API at $2 input/$6 output per million tokens. Third-party benchmarks placed Qwen 3.8 Max fourth in frontend code arena (1,668 Elo), second in vision arena (1,305), and achieved 87.3% on SWE-bench with 66.1 on the Vals Index at roughly 2.3x lower cost-per-test than Claude Opus 4.7. The release signals a strategic shift by Alibaba toward ecosystem influence through open weights, intensifying competition between Chinese and Western frontier models, though licensing restrictions and deployment complexity (requiring 8+ H100/B200 GPUs minimum) limit practical accessibility.

Genspark Open Sources GenOffice: A Free, Ad-Free AI Office Suite for macOS and Windows with Docs, Sheets, Slides, PDF

MarkTechPost 17 hours ago 38

Genspark released GenOffice, an open-source AI-native office suite with document, spreadsheet, presentation, and PDF tools for macOS and Windows under Apache License 2.0. The alpha product at version 0.4.110 consumed approximately $10,000 in API tokens to develop and runs free with no ads, though AI features require a Genspark account and credits. The technical architecture preserves original file bytes by patching only edited content back into source documents, enabling compatibility with Word, Excel and PowerPoint while supporting AI-assisted editing as a core workflow rather than a side feature.

Quoting Steve Yegge

Simon Willison 20 hours ago 7

Steve Yegge described how Gas Town, a reusable system he built, failed when Anthropic's Opus 4.7 introduced a problematic behavior pattern where the model constantly wanted to revise itself rather than converge on a working state. The critical breaking point occurred with version 4.7's release, which Yegge characterized as the final cause of Gas Town's collapse. This failure illustrates challenges in building stable systems dependent on specific LLM behavior patterns and version compatibility.

Notes on the third era of slop

Platformer 20 hours ago 10

A roundup covering an AI notetaker startup rejecting surveillance business models, OpenAI's cyberattack on Hugging Face prompting congressional interest in AI regulation, and arguments for building personal software rather than relying on commercial SaaS platforms. Chris Pedregal's startup Granola refuses to sell user transcript data to employers despite pressure from companies seeking access. The article discusses broader tensions over AI transparency, data ownership, and whether individuals should build their own tools instead of depending on centralized services.

New ways to learn and teach with ChatGPT Work and Codex

OpenAI Blog 21 hours ago 31

OpenAI released new education plugins for ChatGPT Work and Codex designed to support K-12 teachers, college educators, and students in learning, teaching, research, and building activities. The plugins enable integration of AI tools directly into educational workflows without specified technical details or rollout timeline. Teachers and students gain access to AI-assisted capabilities for classroom instruction and project development.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.