Hi there, this is your daily βοΈ AIpresso.
In today's AIpresso:
π€ Salesforce, Nvidia launch CRM AI model
π AI coding agent fails 60% of tasks
β‘ rack cuts AI inference cost 67x
π₯οΈ Claude Code tops rivals at web design
π§ Tool runs CUDA apps on AMD GPUs
Plus: π‘ 5 strategies & tactics, π 6 other news you might like, π§° 6 tools, and π 5 papers.
π€ Salesforce, Nvidia launch CRM AI model LINK
π AI coding agent fails 60% of tasks LINK
β‘ rack cuts AI inference cost 67x LINK
π₯οΈ Claude Code tops rivals at web design LINK
π§ Tool runs CUDA apps on AMD GPUs LINK
π‘ Strategies & Tactics
> Accelerating Dropless MoE Training in JAX with NVIDIA Transformer Engine: Specialized GPU kernels that handle uneven token loads deliver a 10x speedup for mixture-of-experts AI training without sacrificing model quality.
> One coordinate breaks abliteration on Gemma-3: A single dominant activation channel wrecks Gemma-3's refusal-direction editing, and clipping or masking that outlier restores usable results, showing scale outliers can silently break interpretability methods.
> Anthropic analyzed 400,000 Claude Code sessions, and it turns out there's a right way to use it: Set the goal and constraints, then let Claude Code choose how to implement it, because delegating rather than micromanaging gets far better results.
> LLMs as a Judge: How to Know if Your LLM is Healthy: Verify AI agents with automated tests, a second model that scores answers, and human review before launch, since one right reply can't prove reliability.
> Astra appears to perform belief-propagation-like inference without CoT: Testing showed that OpenAI's GPT-6 Astra solves complex logic puzzles without step-by-step reasoning, seemingly weighing probabilities internally and improving with more filler tokens.
Other news you might like
- Nvidia, Palantir, and others restrict advanced AI model usage over privacy concerns, report claims β 'paranoia' rising over customer intellectual propertyLINK
- Trump rejects AI guardrails; for Anthropic, $13.7 billion buys Georgia computeLINK
- Perplexity Portable Computer Is Now Available on Windows, Powered by NVIDIA RTXLINK
- Anthropic Deploys Claude to Automate Financial Adviser Prep WorkLINK
- Chinese AI models dominate OpenRouterβs US token consumption. It can now guarantee that traffic stays entirely in the US.LINK
- ElevenLabs MCP can now generate voice, music, images & videoLINK
π§° Trending tools
Typewise Nova: an AI agent platform that handles customer service, sales, and billing requests end to end, with human oversight for approvals and handoffs.LINK
ChatGPT Images 2.5: generates and edits images from text or sketches, preserving reference-photo subjects and rendering natural lighting up to 50% faster.LINK
Diiverge: turns any image into a branching point-and-click adventure, generating new scenes and short clips as you explore and share saved paths.LINK
Frigade Assist API: embeds an AI agent and React components that walk users through in-app workflows, cutting onboarding and support effort without custom development.LINK
Cognition's SWE-2: a coding model that matches top performers on FrontierCode benchmarks while cutting cost and turns, available in Devin Desktop and CLI.LINK
Perplexity Hybrid Compute: splits tasks between cloud and local models, keeping sensitive files on your Mac while frontier models handle research and drafting.LINK
π Trending papers & reports
AI agent orchestration adds a decision layer that weighs each action's expected payoff against its cost, cutting wasteful tool calls and delays so agents stay useful without running up unnecessary compute bills.LINK
Domain-tuned model repair restores the broad skills a specialized model loses during fine-tuning, using ordinary prompts instead of the original training data, and beats existing methods across role-play and medical question-answering tasks.LINK
Answer-first prompting tells a chatbot to give its answer before explaining why, boosting accuracy in ~88% of tests by an average ~17% while cutting the delay and cost of lengthy reasoning.LINK
Task-specific coaching notes teach an AI agent to write custom instructions for each new job instead of reusing old examples, lifting task completion from ~73% to ~81% on a software benchmark.LINK
Minecraft AI agents get a memory that links what they see to why it matters, letting them recover from failures and reorder tasks on the fly, sharply boosting success on long, multi-step missions.LINK
See you tomorrow for a new dose of βοΈ AIpresso!