Artificial Intelligence

Build an Agentic Event Venue Operator with MongoDB Atlas, Voyage, and LangGraph
Build an Agentic Event Venue Operator with MongoDB Atlas, Voyage, and LangGraph Introduction
This tutorial starts where most agent demos stop: giving the agent persistent memory, operational context, and a place to write back what happened. An event operator does not just need an agent that can summarize a weather report or generate
Zyphra Releases ZUNA1.1: An Apache 2.0 EEG Foundation Model With Variable-Length Inputs From 0.5 To… This week, Zyphra released ZUNA1.1 under the Apache 2.0 license. The EEG foundation model reconstructs, denoises, and upsamples data across arbitrary channel layouts. It builds on ZUNA1, the Zyphra’s earlier open EEG foundation model.
The main change is flexibility, not a
NVIDIA AI Releases Nemotron 3 Embed: An Open Embedding Collection Whose 8B Checkpoint Ranks #1… Embedding models decide which passages an agent ever sees. NVIDIA released Nemotron 3 Embed model to work on that layer. It targets production-scale RAG, agentic retrieval, code retrieval, and agent memory.
What is Nemotron 3 Embed?
The model collection includes three open
AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation arXiv:2607.13230v1 Announce Type: new
Abstract: Agentic AI introduces new insurance challenges because autonomous AI systems can make decisions, invoke tools, modify external environments, and interact with third-party services. This paper develops an AI-native mathematical framework for underwriting, pricing, and contract
Interventional Grounding Audits: Black-Box Premise-Dependency Tests for LLM Chain-of-Thought via Predicate Substitution arXiv:2607.13069v1 Announce Type: new
Abstract: Large language models produce chain-of-thought (CoT) reasoning that appears logically sound yet may not genuinely depend on its stated premises. We introduce interventional grounding audits, a black-box, step-level test of premise dependency: we intervene on
Networked Intelligence: Active Shared Context Graphs for Human-AI Team Science arXiv:2607.13220v2 Announce Type: new
Abstract: Most AI-for-science systems focus on scaling a single reasoning process by using better models, larger context windows, long-horizon agentic execution, or digital co-scientists working with one principal user. However, challenging scientific problems are rarely solved
Set-shifting Behavioral Test for Harnessed Agents arXiv:2607.13396v1 Announce Type: new
Abstract: What happens to an LLM agent’s tool choice when the reliable tool silently changes within an ongoing session? We borrow set-shifting from cognitive psychology to study how well agents adapt to hidden reliability shifts. Our
Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases arXiv:2607.13292v1 Announce Type: new
Abstract: Autoformalization translates informal natural language into formal, machine-verifiable languages. While most work focuses on individual statements, real formalization efforts are inherently theory-level: they require an entire web of axioms, definitions, and lemmas before target theorems
Self-Improvements in Modern Agentic Systems: A Survey arXiv:2607.13104v1 Announce Type: new
Abstract: Self-improving autonomous agents are moving from research prototypes to deployed systems. The primary goal is controllable evolution, or adaptation, from experience with minimal or even no human input. This survey frames modern self-improving agents as
SPINE: Bridging the Cyber-Physical Gap with Agentic AI arXiv:2607.13049v1 Announce Type: new
Abstract: Foundation models have given robots a sophisticated brain for complex decision-making, yet deploying that intelligence into a physical platform still demands tedious, expert-driven calibration. This deployment gap, the robot’s spinal cord, remains a primary bottleneck
Moonshot AI Releases Kimi K3: A 2.8 Trillion Parameter Open MoE Model With Kimi Delta...
Moonshot AI Releases Kimi K3: A 2.8 Trillion Parameter Open MoE Model With Kimi Delta… Moonshot AI just released Kimi K3. It is a 2.8-trillion-parameter model with native vision and a 1-million-token context window. Moonshot calls it the world’s first open 3T-class model.
What is Kimi K3?
Kimi K3 is a sparse Mixture-of-Experts (MoE) model built
OpenAI Details GPT-Red: An Internal Automated Red-Teaming Model That Beat Human Red-Teamers 84% To 13%… This week, OpenAI published details of GPT-Red, an internal-only automated red-teaming model. Its job is to attack OpenAI’s own models and find prompt injection vulnerabilities.
OpenAI gives two reasons. Human red-teaming is time-intensive and does not scale. Commonly used robustness
Create, edit and star in videos with two Google Vids updates
AI
Create, edit and star in videos with two Google Vids updates Gemini Omni and personal avatars in Google Vids make video creation easier than ever.
Connect more of your apps to Search
AI
Connect more of your apps to Search You’ll be able to securely link and interact with your go-to services directly in AI Mode.
Patter SDK Guide to Building a Restaurant Booking Phone Agent with Dynamic Variables, Guardrails, Latency...
Patter SDK Guide to Building a Restaurant Booking Phone Agent with Dynamic Variables, Guardrails, Latency… In this tutorial, we explore the Patter SDK by building a voice-agent workflow that simulates how an AI phone assistant behaves during real conversations. We work with a restaurant booking use case in which we define dynamic caller variables, register
SpaceXAI Open-Sources Grok Build: The Rust Agent Harness, TUI, and Tool Layer Behind Its Coding… SpaceXAI has open-sourced Grok Build, the terminal-based AI coding agent behind its grok CLI. The source landed on GitHub today. The release covers the agent harness, TUI, CLI shell, and developer tooling under the Apache 2.0 license
What is Grok Build?
A
Thinking Machines Lab Releases Inkling: A 975B-Parameter Open-Weights Multimodal MoE With 41B Active Parameters And… Thinking Machines Lab just released Inkling, their first model trained from scratch, weights are open, fine-tunable on Tinker. The lab pitches it as a base for customization.
What is Inkling?
Inkling is a Mixture-of-Experts transformer with 975B total parameters and 41B active.
Soofi Consortium Releases Soofi S 30B-A3B: An Open Hybrid Mamba-Transformer MoE Foundation Model For German… A German research consortium has published the pretraining report for Soofi S 30B-A3B. It is an open base model for German and English. Training ran end to end on Deutsche Telekom’s Industrial AI Cloud in Munich. Preview weights are on
Building a Gin Config Controlled PyTorch Pipeline with Configurable MLP Variants, Cosine Scheduling, and Runtime… In this tutorial, we implement a Gin Config–controlled PyTorch experiment pipeline in which the executable training code remains stable. At the same time, the experimental degrees of freedom are moved into declarative configuration files. We construct a nonlinear spiral binary
Celebrating 25 years of visual search innovation
AI
Celebrating 25 years of visual search innovation Google Images is turning 25. Here’s a look back at some major milestones — and new ways to explore and create visual content.

Load More