Hi there, this is your daily โ๏ธ AIpresso.
In today's AIpresso:
๐ฐ Nvidia backs OpenAI with $105B
๐ Amazon destroys rare books to train AI
๐ LiteLLM hack exposed 2,500 organizations
๐ฅ๏ธ Qwen model runs coding agents locally
๐ป Cursor launches GitHub rival Origin
Plus: ๐ก 4 strategies & tactics, ๐ 7 other news you might like, ๐งฐ 6 tools, and ๐ 5 papers.
๐ฐ Nvidia backs OpenAI with $105B LINK
๐ Amazon destroys rare books to train AI LINK
๐ LiteLLM hack exposed 2,500 organizations LINK
๐ฅ๏ธ Qwen model runs coding agents locally LINK
๐ป Cursor launches GitHub rival Origin LINK
๐ก Strategies & Tactics
> Developing Nemotron 3.5 Lightning NVFP4 with QAD Using NVIDIA Model Optimizer: Compressing NVIDIA's Nemotron model with quantization-aware distillation shrinks it from 66 GB to 22 GB and quadruples throughput while keeping accuracy.
> DeepSeek V4 Pro 0813 vs GPT-5.6 Sol on DeepSWE: Cost, Coding, and Routing: Run cheap DeepSeek first and escalate to pricier GPT-5.6 Sol only when tests fail, hitting higher accuracy for under half Sol's cost.
> I ditched Ollama as my default runtime, and the replacement starts models in a fraction of the time: Switching to BaseRT, a Mac-only runtime for running AI models locally, loads models and processes prompts far faster than Ollama on Apple Silicon.
> The economics of a voice agent: what a minute of conversation actually costs: Judge a voice agent by its cost per successful outcome, not its per-minute rate, since failures, retries, and human handoffs quietly inflate the real price.
Other news you might like
- Stripe will reportedly acquire AI gateway startup OpenRouter for $7B+LINK
- Multi-Vector (Late Interaction) Embedding Models with Sentence TransformersLINK
- Same Cluster, 33 Points More Utilization: What Changed Was the OrderLINK
- U.S. CISA adds a Ray-Project Ray flaw to its Known Exploited Vulnerabilities catalogLINK
- Snowflake targets AI costs with model routingLINK
- Microsoft Copilot reveals secret input that allowed it to be hackedLINK
- Claude can now delete your production voice agent from a chat windowLINK
๐งฐ Trending tools
Zetik: a personal intelligence agent that continuously tracks news, podcasts, papers, and code, delivering concise briefings through feeds, push, newsletters, or RSS.LINK
Hoplite: moves your local coding agent setup, sessions, MCP servers, dependencies, and CLIs, to the cloud so agents run in parallel without laptop interruptions.LINK
Basedash Tasks: builds dashboards and analyzes customer data from plain-language descriptions, so you can skip writing SQL queries manually.LINK
Deskcommcrm: Open-source AI sales OS, self-hosted CRM with native AI agents + WhatsApp (WAHA). Open alternative to Kommo, Octadesk & Intercom for any business that sells by chat. MCP-ready, multi-tenant, LGPD.LINK
world-intel-mcp: a server providing 100+ tools for real-time global intelligence across markets, finance, conflict, cyber, and climate domains.LINK
LLMVault: an intentionally vulnerable training platform for practicing AI security, prompt injection, RAG security, agent security, and GenAI penetration testing.LINK
๐ Trending papers & reports
The ambiguity curse shows that language models struggle most to predict situations with many plausible next words, meaning you should trust their probability estimates least exactly when outcomes are genuinely uncertain and open-ended.LINK
Live-shopping avatar agents can be trained to follow rapidly changing product and compliance rules without retraining, letting a compact model outscore a top general system on real streaming Q&A, ~95 versus ~93.LINK
A model's sense of the current year comes from two disconnected internal clocks, one tied to verb tense and one to direct questions, so no fix updates both at once, undermining reliable time-based reasoning.LINK
Database automation code gets its first real exam, and tests of eight AI systems show they still stumble on writing working procedural database programs, a task ordinary coding and query benchmarks never checked.LINK
Hallucination pinpointing flags the exact words an AI made up and traces each back to the source text it should have relied on, showing not just whether an answer is wrong but where.LINK
See you tomorrow for a new dose of โ๏ธ AIpresso!