Hi there, this is your daily โ๏ธ AIpresso.
In today's AIpresso:
๐ SpaceX buys Cursor for $60B
๐จ๐ณ Alibaba's Qwen leads AI downloads
๐ Naming error let AI attack real firm
๐ค AI models are most confident when wrong
๐ง Google: AI brake alters its beliefs
Plus: ๐ก 5 strategies & tactics, ๐ 6 other news you might like, ๐งฐ 6 tools, and ๐ 5 papers.
๐ SpaceX buys Cursor for $60B LINK
๐จ๐ณ Alibaba's Qwen leads AI downloads LINK
๐ Naming error let AI attack real firm LINK
๐ค AI models are most confident when wrong LINK
๐ง Google: AI brake alters its beliefs LINK
๐ก Strategies & Tactics
> Cutting RAG inference costs 6x starts with deciding what never reaches the LLM: Resolve easy cases with fixed rules first and reserve the language model for the genuinely ambiguous minority, cutting inference costs sixfold while keeping decisions auditable.
> Don't classify. Hallucinate!: Let AI freely invent tags for content, then match those guesses to your real tag list using vector similarity, sidestepping the need to feed it every existing tag.
> I'm running a 284-billion-parameter model across two machines, and it finally matches the cloud: Splitting a giant DeepSeek AI model across two Nvidia desktop boxes matches cloud quality and triples speed, but the cloud API stays far cheaper.
> Tokenmaxxing: Why AI consumption needs control: Have finance teams track AI spending against performance metrics per team so companies can tell valuable use from expensive waste.
> What if Parameter Updates were Text?: Instead of reinforcement learning, this method optimizes readable "advice" text and then bakes it into a model's weights, so humans can inspect how each update shapes behavior.
Other news you might like
- React for Agents: Astro Creator Brings Hooks to his Meta-Harness, FlueLINK
- Optima tackles AI benchmarking's biggest flaw by letting users test models against their own dataLINK
- New benchmark confirms AI models still perform poorly at visual perceptionLINK
- How To Catch a Distilled ModelLINK
- Excel's Copilot function is headed for the Recycle BinLINK
- Anthropic's bio-weapons filter was down for nearly a year, exposing 133 million requestsLINK
๐งฐ Trending tools
Attyn: an AI-powered cursor tool that rewrites text, transcribes speech, explains on-screen content, and visualizes answers directly in your apps.LINK
Clears: an agentic execution platform that automates software delivery tasks across the SDLC, reducing manual coordination between AI tools and workflows.LINK
Vendo: an embedded layer for SaaS products that lets end users build custom views, micro-apps, and integrations using natural language, on your API.LINK
HarnessRouter Community Edition: a unified Agent API that connects Codex, Claude Code, Hermes, and other agent harnesses, letting teams build agent-powered products without managing separate backends.LINK
Chert: lets you build and deploy conversational iMessage agents for customer service or lead capture, with configurable prompts and CRM integrations like HubSpot, Close, or GoHighLevel.LINK
octo-agent: a self-hosted AI assistant that keeps your models and data local, offering coding help across CLI, web, desktop, and mobile interfaces.LINK
๐ Trending papers & reports
Wireless signal decoding gets a self-improving search method that learns to untangle mixed-up signals from many antennas at once, producing more reliable data for next-gen wireless receivers to correctly recover transmitted bits.LINK
Game world simulation separates tracking character skeletons and movement from painting the visuals, so forcing mismatched actions shifts joint accuracy by ~31%, proving the underlying state, not just the pixels, actually controls what happens, letting long, glitch-free interactive scenes be fixed at the state level instead of the video level.LINK
Ancient hand stencils can now be sexed with a probability score instead of a single guess, using AI models trained on 14,036 modern hand images that hit over 88% accuracy on older age groups, giving archaeologists a defensible, uncertainty-aware read on who made Paleolithic cave art.LINK
Wheat farming data now links nitrogen and disease records from different sources into one searchable system, letting a single question pull combined answers researchers previously had to hunt for across separate datasets.LINK
Japanese riddle solving shows top AI models correctly guess the answer internally but often fail to commit to it, scoring only ~18% versus humans' ~53% on these insight puzzles.LINK
See you tomorrow for a new dose of โ๏ธ AIpresso!