AI News
4 sources Β· 50 articles
Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence
Perplexity Research and turbopuffer have released pplx-embed-v2-context-9b-preview, a contextual embedding model for RAG pipelines. Each chunk is embedded with the full document in view. The real chan
Just nowGoogle DeepMind Unveils Gemini 4 Argon with 1M Output Tokens for Coding, Knowledge Work and Cyber Defense
Gemini 4 Argon tops GPT-6 Astra and Claude Opus 5.5 on most benchmarks, but access remains gated today. The post Google DeepMind Unveils Gemini 4 Argon with 1M Output Tokens for Coding, Knowledge Work
6h agoOpenAI Releases GPT-6.1 Sol: Near-Astra Coding and Computer Use at One-Fifth of Astraβs Token Price
OpenAI released GPT-6.1 Sol on September 29, 2026, an upgrade to GPT-6 Sol. It reaches near-Astra results on agentic coding, computer use and professional work at one-fifth of Astra's token prices. It
7h agoDid AI Just Solve One of Mathematicsβ Biggest Problems?
OpenAIβs agents reached a proposed solution in 88 hours. But the human research that came before, and the controversy that followed, raise a harder question: what actually counts as an AI discovery?
14h agoOllama for Managing Local Language Models: A KDnuggets Cheat Sheet
Ollama pulls model weights, keeps an HTTP server on port 11434, and hands any client an OpenAI-shaped endpoint pointed at your own machine. Learn how to manage, configure, and optimize using Ollama ri
16h agoPerplexity Introduces Photon: A Rust-Based Retrieval Engine That Cuts p99 Latency From 800 ms to 65 ms
Perplexity has released Photon, an in-house retrieval and ranking engine written in Rust. It replaces an open-source engine Perplexity had forked for its AI-native search stack. Photon now handles ret
19h agoNVIDIA Researchers Introduce Physis-Lang: Self-Evolving Physical Language That Lifts Cosmos 3 Past Veo 3.1 on Physics Benchmarks
Video world models can render convincing clips that still break physics. Butter spreads like paint. Balls pass through walls. A team from NVIDIA, MIT and the University of Oxford argues the fix can co
20h agoCrawlRaven MCP
Your SEO work, done from your AI agent Discussion | Link
21h agoBevell
CAD automation where it counts. Discussion | Link
1d agoCoIsland
Your whole engineering stack, in a notch Discussion | Link
1d agoOne Bad Prompt Took Down a Companyβs Salesforce: RSAβs Jim Taylor on Agent ID and Taming the 4,000 Shadow AI Agents Hiding in Your Enterprise
RSA launched Agent ID at The AI Conference in San Francisco. It's an agentic identity security platform for finance, government, healthcare, and critical infrastructure. It has 3 modules. Discover fin
1d agoOpen TTS Leaderboard: Scalable Evaluation for Multilingual Text-to-Speech and Voice Cloning
1d agoLiquid AI Releases d1: A Decision Model That Returns Calibrated Probabilities With Zero Output Tokens
Liquid AI has released d1, a decision model built for structured choices instead of text generation. You give it context and a set of typed questions. It returns calibrated probabilities across a fixe
1d agoOpenAI Launches dots: Always-On GPT-6 Astra Agents That Work From Their Own Cloud Computers
OpenAI just introduced dots at their DevDay today. Dots are persistent AI agents powered by GPT-6 Astra. Each dot gets its own cloud computer and browser. It works across 4,000+ apps through ChatGPT p
1d agoNebius Opens 2026 Physical AI Awards: Five $150K Compute Credit Prizes
Nebius and NVIDIA are running the 2026 Physical AI Awards for startups with products in the field. Five category winners each get $150,000 in compute credits, joint promotion, executive mentorship, an
1d agoHow to Turn Excel Data Into PowerPoint Presentations With AI
Learn how to use Julius AI to analyze Excel data, verify key findings, and turn them into an editable PowerPoint presentation.
1d agoNVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction
1d agoSuperWhisper s1-mini: The 600M Parameter Model Built Just for Transcription
This is a summary of, and insights into, what I found digging into the recently-released Superwhisper S1 family of voice-to-text models.
1d agoEvlat
Know which AI coding agent is waiting on you Discussion | Link
1d agoGetting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents
1d ago5 Free Courses to Learn AI Engineering
Learn LLM fundamentals, AI engineering, RAG, MLOps, fine-tuning, and deployment with five practical courses designed to help you become a stronger AI and machine learning engineer.
1d agoAktar
Share files instantly from your own cloud. Mac, Windows, iOS Discussion | Link
1d agoGoogle Research Open-Sources RRSI: AI Agents That Improve Their Own Harness Without Overfitting
Google Cloud AI Research has open-sourced RRSI, a framework that lets LLM agents rewrite their own prompts, tools and memory while model weights stay frozen. It adds a leakage critic, a noise floor, a
1d agomβkay
One voice for all your coding agents, from your phone Discussion | Link
2d agoAgent Identity
Give every AI agent real identities, inboxes, and a phone Discussion | Link
2d agoGemini 3.5 Transcribe vs OpenAIβs GPT-Transcribe
Here's how each got to where it is, a real use case and working code for both, and a side-by-side on the numbers that actually matter.
2d agoFlocker Agent Profiles
Profile Pages for Agents: your live AI collaboration network Discussion | Link
2d ago3 Numba Tricks for Python Runtime Optimization
When Numba code disappoints, it's nearly never the compiler, and usually ends up being the boundary around the compiled code: not crossing it, not making it wide enough, or crossing it during every ru
2d agoGitBot
Build bots on the coding agent you already use Discussion | Link
2d agoHolo4: powering generalist computer-use agents
2d agoZumbo
Open source local AI voice-to-text for Mac for private STT Discussion | Link
2d agoVoice Memo
Open-source voice notes that file your tasks and reminders Discussion | Link
2d agoSquint
Drag a box on your screen and ask AI about it Discussion | Link
3d agoSpeek
A context aware FOSS voice assistant and dictation for macOS Discussion | Link
3d agoBatching by Length Instead of Looping Item by Item for SLM Optimization
We finish off our short series on SLM optimization with the third entry, focused on batching by length instead of looping item by item.
5d ago7 Advanced Python Tricks to Level Up Your Coding Skills
Leveling up rarely means new syntax. It means learning what the language already promised you.
5d agoRinkata
One source of truth for your team and its AI agents Discussion | Link
5d agoAccelerating vision-language models with LFM2.5-VL-DSpark
6d agoMCP Explained in 5 Minutes
A visual guide to MCP that explains how it works, how to use it with Claude Code, Tavily, GitHub, and Playwright, and what is new through simple diagrams that make the whole concept easy for anyone to
6d agoHow UK AISI and EvalEval Are Making Benchmark Results Reproducible
Sep 22Transformers now runs llama.cpp quants
Sep 22Jun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX community
Sep 22tokenizers v1: encode, decode and scaling, measured
Sep 21Pexo
Produce pitch perfect launch videos with precise control Discussion | Link
Sep 18