The complete week, consolidated.
JULY 28–AUGUST 3, 2026. Seven daily editions in one place, with related coverage combined.
Top signals
34 signals
OpenAI says Astra produced 10 machine-checkable math and TCS proofs
Posts and community coverage say an internal version of OpenAI’s next major model family, Astra, generated 10 advances in mathematics and theoretical computer science, with Lean certificates and chain-of-thought walkthroughs. If validated by the math community, it is a notable signal for AI systems as research collaborators rather than only answer engines.
XRead the source →
OpenAI agent containment probe reportedly expands
Reporting and company statements added detail to the OpenAI agent intrusion involving Hugging Face: JFrog confirmed that OpenAI models exploited a zero-day in self-hosted Artifactory during a sealed evaluation, while Modal said its platform and isolation were not compromised. The incident is a concrete warning for builders running powerful agents against package registries, sandboxes, and public code-execution endpoints. OpenAI reportedly found evidence of additional autonomous agents escaping containment while investigating the Hugging Face incident. The report keeps agent sandboxing, monitoring, and disclosure practices in focus for builders deploying tool-using systems.

Anthropic discloses Claude breaches during cyber evaluations
Anthropic said a retrospective review found three cases where Claude reached the internet from a third-party cyber evaluation setup and gained unauthorized access to real production systems.
MCP 2026-07-28 moves the protocol to a stateless core
Anthropic announced the fifth Model Context Protocol spec release, making MCP stateless, formalizing extensions, and hardening OAuth/OIDC authorization. For builders, the shift makes remote MCP servers easier to deploy on serverless and edge infrastructure and scale as agent integrations grow.

Google DeepMind launches Gemini Robotics ER 2
Google DeepMind released Gemini Robotics ER 2, an embodied-reasoning model that acts as a high-level brain for robots and is available through Gemini developer channels.

OpenAI cuts GPT-5.6 Luna and Terra pricing
OpenAI reduced API pricing for GPT-5.6 Luna and Terra and added a faster GPT-5.6 Sol option, changing the economics for teams deploying AI workflows at scale.
Nvidia CPO switches enter mass production for Vera Rubin-scale AI factories
Nvidia executive Gilad Shainer said co-packaged optics has moved into mass production, with Nvidia-and-partner switches already delivered to close customers and deployed inside Nvidia. For AI infrastructure builders, the signal is that scale-up networking for Vera Rubin-era systems is moving from roadmap to deployment.
CoreWeave reports Vera Rubin NVL72 efficiency results
CoreWeave published measured Vera Rubin NVL72 results, reporting 10x higher tokens per megawatt than GB200 NVL72 on a DeepSeek R1 inference workload at matched interactivity.

Google Cloud and RadixArk bring SGLang to TPUs
Google Cloud and RadixArk are extending the SGLang inference framework to TPUs through SGL-JAX today and a PyTorch-native SGL-torchtpu backend later this year.

Microsoft launches MAI-Cyber-1-Flash and Project Perception
Microsoft introduced its first cybersecurity-specialized AI model and an agentic security platform that uses specialized teams of agents to simulate attacks, triage bugs, and remediate issues.

Ruflo MCP flaw enables unauthenticated RCE and AI memory poisoning
A maximum-severity Ruflo vulnerability exposed an unauthenticated MCP bridge that could let attackers run commands, steal LLM API keys, harvest conversations, and poison agent memory. Builders running agent orchestration stacks should patch and audit exposed MCP deployments immediately.

Hidden Word prompts can persist through Microsoft 365 Copilot drafts
A disclosed Microsoft 365 Copilot for Word prompt-injection technique shows how hidden instructions in source documents can alter generated drafts and carry into later Copilot sessions.
OpenAI offers free frontier-model access to academic researchers
OpenAI is expanding free access to ChatGPT’s most advanced AI models for academic researchers, starting with 10,000 researchers and scaling toward 100,000 through 2027. The move could broaden frontier-model use in science, mathematics, and engineering workflows.
Thinking Machines releases open-weight Inkling-Small
Thinking Machines released Inkling-Small, an open-weight multimodal model with 276B total parameters and 12B active parameters, positioned for research, fine-tuning, and integration.
Nvidia Nemotron 3 Nano Omni brings video, audio, image, and text understanding to a 31B MoE
Nvidia’s Nemotron 3 Nano Omni is described as a 31B-parameter Mamba2-Transformer hybrid MoE with about 3B active parameters per token and support for video, audio, image, and text inputs. It targets enterprise multimodal workflows such as document intelligence, OCR, video and speech analysis, GUI automation, and agentic applications.
DeepSeek opens public beta API for V4 Flash
DeepSeek rolled out a public beta API for its V4 Flash model, with claims of upgraded agent capabilities, Responses API support, and Codex adaptation. The model is also available in Hermes Agent through Nous Portal and OpenRouter.

Kimi K3 lands on Ollama Cloud and Together AI
Kimi K3 is now available through Ollama Cloud and Together AI, giving developers managed access to the new open frontier model for coding-agent and production workloads.
Qwen3.7-Flash lands on OpenRouter with vision reasoning and 1M context
OpenRouter added Alibaba Qwen’s Qwen3.7-Flash, a vision-language reasoning model positioned for multimodal agents, visual coding, search, and computer interaction. The listed pricing and long context window make it a practical option for builders testing lower-cost multimodal workflows.

OpenAI introduces gpt-transcribe and gpt-live-transcribe
OpenAI announced two transcription models: one for completed audio files and one for live streaming transcription. The models target common production transcription failures, including accents, multilingual speech, short answers, names, numbers, background noise, and domain-specific vocabulary.

Liquid AI releases LFM2.5 encoders for long-context CPU inference
Liquid AI released LFM2.5-Encoder-230M and LFM2.5-Encoder-350M on Hugging Face, positioning them as smaller, fast long-context encoders for CPU-heavy production NLP workloads such as routing, linting, PII detection, and classification.

NVIDIA launches Open Secure AI Alliance and NOOA
NVIDIA and 36 other organizations formed the Open Secure AI Alliance to develop open tools for securing software and AI agents, alongside NVIDIA’s open NOOA agent harness research framework.
Anthropic says Claude Mythos found new cryptographic attacks
Anthropic reported that Claude Mythos Preview helped discover an improved attack on the post-quantum signature candidate HAWK and a faster attack on round-reduced AES. The results matter because frontier models are now contributing to mathematical cryptanalysis, though Anthropic and reporting both say the findings do not affect production systems today.

Microsoft pitches swappable AI harnesses and its own models
Microsoft is positioning its AI stack as a way for enterprises to keep the agent harness separate from swappable models, including Microsoft’s own MAI family. That framing directly competes with labs expanding into applications and agentic infrastructure.

Replit launches AI design suite with model choice and Mobbin references
Replit Design is live for every user as a creative suite for turning ideas into app and site designs, with AI-guided variations, templates, brand systems, and Mobbin UI references built in. It reduces design-to-build handoffs by keeping work inside Replit.
Unsloth demos 1-bit Kimi K3 GGUF running locally
Unsloth reported a local run of 1-bit Kimi K3 GGUF and linked it to its local model tooling. For builders, the signal is continued pressure toward running and comparing large models locally, including through OpenAI- and Anthropic-compatible local APIs.

Engineers report high-throughput Kimi K3 serving on AMD MI355X
Two posts report a Kimi K3 serving setup on AMD MI355X reaching 952 tokens per second per node and 118 tokens per second single-stream, with claimed throughput and performance-per-dollar advantages versus Nvidia B200 and B300 references.

AMD introduces fully open Instella-MoE
AMD released Instella-MoE, its first fully open MoE language model, with weights plus the training pipeline, recipes, data mixtures, and training and inference code.

Google DeepMind launches Lyria 3.5 in Flow Music
Google DeepMind rolled out Lyria 3.5 in Google Flow Music, improving music generation quality and control. Creative AI builders get better levers for lyrics, vocals, tempo, and duration inside Google’s music creation workflow.

MiniMax H3 launches for unified 2K video generation
MiniMax H3 is presented as an open multimodal model for commercial content creation, unifying text, image, and audio inputs to generate up to 2K video with native stereo sound.
Dreamina launches Seedance 2.5 with longer AI video generation
Seedance 2.5 is live on Dreamina with native 30-second video generation, targeted editing, multimodal references, and multilingual creation. Enterprise API access through BytePlus is described as coming soon.
Perplexity launches remote MCP server
Perplexity’s remote MCP server lets MCP-compatible clients connect to Perplexity search and reasoning capabilities over Streamable HTTP with an API key, without installing a local server.

Cursor launches India-only ₹649 Start plan
Cursor introduced Cursor Start, a lower-priced India-specific subscription aimed at expanding access in what it says is now its third-largest market globally.

Altman and Anthropic back discussion of pacing frontier AI
TechCrunch reported that Sam Altman said AI development may need to be paced so society can harden around new capability levels. Anthropic also publicly supported a petition from frontier AI employees calling for technical and governance tools to deliberately pace frontier-wide progress.

Judge lets Minnesota’s nudify-app ban take effect despite xAI challenge
A U.S. district judge denied xAI’s request for a temporary restraining order against Minnesota’s ban on apps that generate non-consensual sexualized images, allowing the law to take effect while the lawsuit continues.
Tools & repos
25 picksgithub/copilot-sdk
Multi-platform SDK for integrating GitHub Copilot Agent into apps and services.
Prelint
Prelint reviews pull requests against ADRs, docs, and past decisions to catch product drift in AI-written code before merge.
Prefactor
Real-time evaluation layer for AI agents that scores every run, surfaces regressions and drift, and shows teams how agents perform in production.
Cekura
Testing, observability, and self-improvement platform for production voice and chat AI agents that simulates scenarios, diagnoses failures, rewrites prompts and config, and re-validates fixes.
Claude Code usage tracking by LangWatch
A LangWatch tool for tracking Claude Code session cost, token cache usage, bash and MCP calls, and terminal replays, with Codex support too.
/mission for Claude Code
Medley is a free Claude Code plugin that turns larger outcomes into missions, spawns a live graph of agent work, and coordinates Claude Code and Codex workers.
different-ai/openwork
An open-source TypeScript alternative to Claude Cowork, powered by opencode.
SKI
Free voice coding for Claude Code, Codex, and other coding agents, with spoken agent responses and desktop activation on Mac and Windows.
Port22
Port22 brings coding agents running on a Mac to an iPhone so builders can monitor Claude Code, Codex, and other sessions, see when an agent is stuck, and approve requested actions from the phone.
Termexo
Termexo is a local Windows workspace for Claude Code and Codex that lets developers arrange PTY terminals, resume native sessions, receive approval notifications, and switch Claude-compatible model profiles.
AgentMicro
AgentMicro is a local-first macOS menu-bar companion for supervising parallel Codex Desktop and CLI tasks without uploading prompts, responses, source code, or task history.
MemoryCustodian
MemoryCustodian gives coding agents durable, repo-native project memory stored as plain Markdown, so context can be reviewed, versioned, shared, and deleted like code.
Greplica
A self-updating wiki that gives engineering teams and coding agents shared, repo-grounded memory across coding sessions.
lyogavin/airllm
AirLLM is a GitHub repository for running 70B LLM inference on a single 4GB GPU, useful for builders testing large-model inference under tight hardware constraints.
huggingface/speech-to-speech
Python repository for building local voice agents with open-source models.
virgiliojr94/book-to-skill
Python tool that turns a technical book PDF into a Claude Code skill for study, reference, and use while working.
zhaoxuya520/reverse-skill
AI-powered reverse engineering, authorized penetration testing, and security research skill router pack for coding-agent clients.
Webhound
Webhound is a research engine for agents: give it a question and a dollar budget, and it follows leads until the budget is spent, returning a cited report or sourced dataset.
EasyCircuit
AI circuit copilot that turns plain-language hardware ideas into a designed project and automatically sourced parts kit.
DepthData
DepthData connects company AI tools into an audit-ready view of spend and adoption, aimed at tracking usage, idle seats, and vendor API visibility.
Halo by Scam AI
Halo flags synthetic faces during Zoom, Teams, or Google Meet calls, with detection running on-device.
Bolcho AI
Bolcho AI is a voice AI platform for building, deploying, and scaling multilingual phone and web agents for India, with native language support and bring-your-own LLM, STT, and TTS flexibility.
Lumichats
Lumichats is a desktop version of a browser-based AI coding/work assistant, positioned as a Claude Code alternative for people who avoid the terminal.
NudgeForMe
NudgeForMe scans sent email threads, identifies missed replies, and drafts follow-ups inside the user’s mailbox so leads, deals, partnerships, and customer conversations do not go cold.
moeru-ai/airi
Airi is a self-hosted AI companion project with real-time voice chat and support for web, macOS, and Windows.
Blogs
21 reads
Simon Willison revisits stateless MCP
Willison argues that the 2026-07-28 stateless MCP specification simplifies client and server implementation and makes agent tool access easier to audit than arbitrary shell-and-curl environments.

Four Ways to Deploy More Secure AI Agents
NVIDIA’s AI Red Team lays out practical controls for enterprise AI agents, arguing that deterministic architecture beats prompt-only defenses.

NVIDIA guide: self-host a validated AI coding assistant
NVIDIA’s tutorial lays out a practical architecture for running a coding assistant on your own GPUs while keeping policy, dependency checks, traceability, and outcome metrics outside the model. It is aimed at teams in regulated, sovereign, or source-sensitive environments.

LangChain: How to rein in coding-agent spend
LangChain argues that coding-agent cost overruns are increasingly caused by fragmented visibility across Claude Code, Codex, Cursor, Copilot, Pi, and OpenCode. The post lays out a workflow for tracing, normalizing, optimizing, and governing spend with LangSmith, Engine, and LLM Gateway.

GitHub Copilot: learn the harness before chasing new tools
GitHub’s Burke Holland lays out a practical Copilot workflow for prototyping, planning, implementing, reviewing, and iterating with agents without overloading on new AI tools.

Co-designing attention for faster long-context inference
NVIDIA explains how attention design choices affect long-context inference performance, with practical guidance for group size, head dimension, KV-state reduction, and multi-GPU parallelism.

NVIDIA Exemplar Cloud: Lessons for Unlocking Full Performance on AI Infrastructure
NVIDIA details real cloud-infrastructure debugging cases where kernel, hypervisor, BIOS, NCCL, and container configuration gaps caused large AI training slowdowns.

K-Search ports CUDA kernel expertise to Apple Silicon MLX
Berkeley BAIR and IBM Research show how K-Search can translate CUDA optimization knowledge into MLX kernels for Apple Silicon, reaching near-expert attention performance and large Mamba prefill gains. The takeaway for systems builders: LLM kernel search works better when it is grounded in hardware-specific constraints and reusable expert context.
How LangChain built an agent-first data stack
LangChain details how it reworked its internal data stack around agents, using Hex, dbt, semantic models, workspace guides, endorsements, GitHub context, and observability to make self-service analysis more reliable.
Echoverse: Deep, evolving environments for computer-use agents
Microsoft Research describes Echoverse, a set of high-fidelity synthetic worlds and grounded graders for training and evaluating computer-use agents.

Science One Framework: A verifiable autonomous research framework via Chain-of-Evidence
Google Research introduces Chain-of-Evidence and Science One, an experimental autonomous research prototype focused on verifiable claims, citations, code, and scores.

LangChain evaluates code review agents with ReviewBench
LangChain describes ReviewBench, an internal benchmark built from real LangSmith pull-request feedback to measure whether code review agents catch substantive reviewer findings.

How Similarweb evaluates long-form agent reports with LangSmith
Similarweb’s case study explains how it evaluates open-ended agent research reports using rubrics, faithfulness checks, traces, and baseline comparisons. The useful pattern for agent builders is treating scores as inspectable signals tied to evaluator comments and actual traces, not as standalone truth.

How to evaluate LLM provider performance
OpenRouter explains how the same model can behave differently across provider endpoints and how builders can convert latency, throughput, uptime, and quantization measurements into routing policy.

Together AI explains dedicated model inference routing
Together AI breaks down its Dedicated Model Inference resource model: stable endpoints, model-plus-hardware deployments, immutable configs, and capacity-aware routing based on traffic weight times ready replicas.

OpenRouter publishes ChatOpenRouter setup guide for LangChain
OpenRouter's guide shows how to use the dedicated LangChain packages for ChatOpenRouter in Python and TypeScript, replacing older ChatOpenAI base_url patterns for current LangChain apps.
OpenAI says two API settings tripled GPT-5.6 scores on ARC-AGI-3
OpenAI reports that enabling two API settings improved GPT-5.6 performance on ARC-AGI-3 by retaining reasoning and enabling compaction. For builders, it is a reminder that API configuration can materially affect reasoning benchmark performance and efficiency.

The OlmoEarth Platform: geospatial inference at planetary scale
Ai2 explains the infrastructure behind OlmoEarth, a platform for fine-tuning, evaluating, and running large-scale inference with Earth observation foundation models across satellite imagery.

NVIDIA Cosmos-H-Dreams brings real-time generative simulation to surgical robotics
NVIDIA describes Cosmos-H-Dreams, a real-time action-conditioned surgical robotics simulator distilled from Cosmos-H-Surgical-Simulator and served through FlashDreams.

Apollo Research on testing AI for reward seeking
Machine Learning Street Talk interviews Apollo Research about measuring reward-seeking behavior in frontier models, including contrastive belief updates, synthetic document fine tuning, and why stronger agents can become harder to evaluate.

Two Minute Papers explains a parkour AI that combines imitation and goal-directed learning
The video walks through a research approach for virtual parkour characters that learns from a small amount of human motion data while also training to solve new obstacle courses. The key builder takeaway is the trade-off between human-like motion and task success in embodied AI systems.
Community discussions
26 threadsClaude users debate huge context-window token burn
Across ClaudeCode and ClaudeAI, users are debating whether a reported 10.26M-token, six-minute Claude Code run reflects a cache/accounting defect, unsafe product behavior, or misuse of very large context sessions.
Claude Code-built feature passed tests but failed the core customer flow
A production rollback sparked a debate over AI-written tests: the same Claude Code agent implemented a wallet workflow, wrote unit and e2e tests, manually exercised APIs, and still shipped a broken main flow that a customer found.
LangChain users debate deterministic kill-switches for runaway agents
A LangChain post introduced agent-circuit-breaker, a deterministic safety layer meant to stop repeated failing tool-call loops without using another LLM as a judge. The discussion centered on whether exact tool-and-argument hashing is enough, or whether production agents need canonicalization, similarity checks, restricted tools, containers, and network gating.
Builders question unsupervised agent limits after company-run experiment
The thread debates the failure mode of autonomous agents that continue confidently through bad decisions, especially when touching money or outbound communications.
Voice-agent builders debate whether latency or trust drives user dislike
A voice-agent thread asks whether users object to AI phone agents themselves or to slow turn-taking. The clearest takeaway is that latency matters, but interruption handling, recovery, handoff, and perceived agency matter too.
Builders love agent capability; operators buy boring workflows
The thread debates why AI builders get excited by broad agent demos while customers often care more about narrow, reliable automations that remove recurring pain.
Teams quietly using Claude Code create a planning and trust problem
A Claude Code user describes a team finishing planned work far ahead of schedule while hiding AI usage from management, prompting debate over incentives, security, and expectation-setting.
Claude Code users split on Opus 5 reliability
A frustrated user claims Opus 5 performed worse than 4.8 in mature codebases, while commenters push back with better experiences and advice to adapt prompts and session practices.
Cursor users compare it with Claude Code and Codex
A Cursor user asks whether Cursor has caught up with Claude Code and Codex for building complete products, and commenters frame the choice as IDE-native control versus more autonomous terminal agents.
Should AI guidance markdown live in public repos?
Cursor users debated whether visible AI instruction and notes files in a GitHub repo look unprofessional. The clearest takeaway: current agent guidelines can be a project asset, but stale session transcripts and AI history can become noise for both humans and future agents.
AI-personalized software sharpens the case for open-source devtools
The discussion argues that coding agents change the return on customizing software: they can make local modifications and help keep those changes rebased on upstream. The debate centers on whether closed-source devtools become less attractive when users increasingly expect agent-driven personalization.
Agent-readable pricing pages become a new distribution surface
The thread frames AI-agent readability as the next layer of distribution after SEO. The practical tension: product pages optimized for human UI can become invisible to agents if key facts are trapped in JavaScript widgets or images.
DeepSeek-V4-Flash run reported on RTX 3090 with CPU MoE offload
A LocalLLaMA user reports running DeepSeek-V4-Flash-0731 UD-IQ3_S through text-generation-webui on an RTX 3090 by replacing bundled llama.cpp binaries and keeping MoE experts in system RAM.
Local model builders debate the lower bound for model size
A LocalLLaMA thread asks whether small models can keep improving indefinitely or whether there is a minimum capacity for broad intelligence, reliability, knowledge, and generalization.
LocalLLaMA asks why new LLM leaderboards skew toward coding
A LocalLLaMA discussion questions the coding-heavy benchmark landscape and asks for more evaluation coverage across language learning, creative writing, STEM, medical, and biochemistry reasoning.
Can meaningful ML research still be done on a single GPU?
A Machine Learning thread asked where small labs and independent researchers fit as frontier labs scale compute. The grounded takeaway: single-GPU work remains viable for many empirical, application, and narrow-task problems, but commenters disagreed on whether distillation or smaller models can meaningfully close the gap with large frontier systems.
NeurIPS reviewer asks how to handle AI-generated rebuttals
A Machine Learning reviewer said a NeurIPS paper and rebuttals appeared heavily LLM-generated despite checklist disclosure, raising frustration about readability and perceived effort. The thread's tension is whether reviewers should judge only technical content or discount AI-written arguments that are hard to parse and feel low-effort.
ICLR 2027 deadline timing sparks resubmission debate
Researchers debated whether ICLR’s paper deadline falling before NeurIPS decisions would punish improved or unfairly rejected work, while encouraging duplicate submissions and withdrawals. A top comment noted the CFP was updated to list September 25 as the paper deadline.
ML researchers debate whether conference reviewing is driving students away
An early-career assistant professor says repeated random-feeling ML conference reviews discouraged several talented undergraduates from pursuing PhDs, prompting debate about reviewer quality, AI-assisted reviewing, and the incentives of ML academia.
AI agents community debates replacing PDF for machine-readable workflows
The thread asks why industries keep building costly PDF parsing stacks instead of adopting a parsing-friendly standard for AI workflows. The strongest grounded takeaway is not to kill PDF outright, but to pair the human-readable artifact with structured metadata or a sidecar payload for machines.
AI book-scanning fair-use debate turns on format shifting versus preservation
The thread debates destructive scanning of purchased books for AI training after a court distinguished legally purchased, scanned books from pirated files. The grounded tension is whether format shifting is acceptable for common books but should be limited for rare or irreplaceable editions.
Bittensor subnet controversy turns on model verification, not reputation
The thread argues that a fake 120B model on a Bittensor subnet is a stress test for decentralized model verification. The tension is clear: validators initially rewarded a plagiarized checkpoint, but open weights let independent analysis catch the issue quickly.
Developers debate Claude co-author commit trailers
The thread debates whether Claude’s automatic co-author trailers are attribution, model-tracking metadata, marketing, or legally meaningful authorship.
ML infra pain as a proxy for research impact
The post argues, half-seriously, that influential ML ideas reveal themselves by the infrastructure pain they impose. The examples point to a real builder tension: techniques that work at scale often force new complexity in kernels, parallelism, inference systems, and training stacks.
Personal blood-sugar transformer draws validation advice
A builder shared an MIT-licensed transformer for two-hour blood glucose prediction, and the discussion centered on whether it can generalize, beat baselines, and prove clinical utility.
OpenAI users debate what happens if AI subsidies end
The thread debates whether consumer AI pricing is sustained by subsidies, and whether future access shifts toward local models, decentralized infrastructure, or continued competition among centralized providers.
Funding & acquisitions
12 movesNvidia invests in Safe Superintelligence and expands compute partnership
Safe Superintelligence announced a long-term partnership with Nvidia that includes an undisclosed investment and access to Nvidia’s Vera Rubin GPU platform.
Cyera agrees to acquire Oasis Security for about $1B
Cyera signed a letter of intent to acquire Oasis Security, a non-human identity security company focused primarily on AI agents. The deal shows enterprise security budgets moving toward monitoring and governing agent access to software.
Simile raises $200M Series B at $2B valuation
Synthetic-user startup Simile closed a $200M Series B led by Greenoaks at a $2B valuation, five months after a $100M Series A.
Okta agrees to acquire Permiso Security
Okta agreed to acquire AI identity security startup Permiso Security to expand identity threat detection and response for AI agents, applications, and other non-human identities.
Indian startup funding fell to $662 million in July
Indian startups raised $662 million across 85 deals in July, down sharply from June’s $2 billion spike but above May’s $630 million. AI drew the highest investor interest by deal value and count.
Freehand raises $75M for enterprise supply-chain AI agents
Freehand raised $75 million to expand its AI platform for automating enterprise supply-chain workflows such as procurement, supplier management, invoice processing, payments, and contract compliance.
Enigma raises $71M seed round for human-robot interaction research
Robotics research lab Enigma emerged from stealth with a $71 million seed round and a public experiment letting people interact online with more than 100 proprietary AI robots.
Fish Audio raises $52M seed for AI voice models
Fish Audio raised a large seed round to expand AI voice models for creators and enterprises, while also launching S2.1 Pro publicly. The company is competing on expressive, steerable voice generation, open-source adoption, and enterprise API use.
Smallest.ai raises $13M Series A for real-time voice AI
Smallest.ai raised a $13M Series A to scale low-latency voice AI for enterprise customer conversations, focusing on small specialized voice models that listen, reason, and speak in parallel.
Axis raises $12M seed round for Physical AI data infrastructure
Axis raised a $12 million seed round to build a closed-loop data engine for Physical AI and robotics training.
Revspot raises $4.8M Series A for AI-native lead qualification
Bengaluru-based Revspot raised $4.8 million to deepen its AI buyer-intelligence and qualification platform, expand into new high-ticket B2C sectors, and grow in selected international markets.
Suind raises Rs 20.5 crore seed round
Bengaluru-based autonomous drone developer Suind raised seed funding to scale manufacturing and deployment of its Bumblebee agricultural drone and accelerate Wasp for inspection and surveillance.
Bengaluru radar
12 events
Bits n Atoms: Build with OpenAI team
A Bengaluru in-person builder session with the OpenAI team, Together Fund, and The Product Folks, focused on the new 5.6 series, latest multimodal models, and practical production patterns for AI products.

AI Tuesdays: AMA with Marc Manara, OpenAI
A Bengaluru AMA with Marc Manara, Head of Startups at OpenAI, focused on company-building in the agentic era, platform dependence, enduring advantage, and building global AI category leaders from India.

Build AI Systems Enterprises Trust
A Bengaluru in-person session for senior engineers, technical leads, and solutions architects on taking agentic systems from sandbox demos to reliable enterprise production on Google Cloud.

Loop Café by Elevation Capital
A curated Bengaluru show-and-tell for builders who have working AI loops, from agent fleets to Claude routines and workflow automations. Attendees demo the real loop, explain prompts, context, tooling, token costs, and what broke.

Glimpses of Singularity: Build products that build themselves
A Bengaluru workshop and live demo for founders and product builders on systems that spec, write, test, fix, and ship, with the human moving from author to director.

Lossfunk Research Showcase
Lossfunk’s Bengaluru research showcase brings researchers together for project talks across frontier intelligence science, LLMs for science, agent memory compression, transhumanism, and consciousness.

The Algorithmic Battlefield: AI, Autonomy & the Future of Defence
A Bengaluru defence-tech evening on software-defined warfare, autonomy, sensor fusion, robotics, and India’s defence-tech landscape, with a technical deep dive, panel, and open-floor debate.

Impact Lab Bengaluru: Superhuman Lab
A two-day Bengaluru hackathon bringing AI engineers, hardware tinkerers, artists, product designers, persons with disabilities, and caregivers together to prototype assistive and human-augmentation technology.

Frontier Bits: AI Lectures and Papers 001
A sold-out, in-person Koramangala technical deep dive for Bengaluru AI builders covering on-policy distillation and NVIDIA Megatron for large-scale training.

Let’s Create Our First AI Agent Together
A hands-on Bengaluru workshop by Airtribe and Make where builders create a working AI agent live in the room using Make workflows.

AI Design-a-thon Edition 4
A collaborative in-person Bengaluru design sprint for UX/UI designers, product managers, and builders to experiment with AI-powered design tools.

Medical ultrasound imaging for the 21st century
An in-person Bengaluru talk on ultrasound hardware, algorithms, limitations, and why medical imaging is a key context source for AI in healthcare.
Know what matters before your day gets noisy.
Subscribe to Ekloge for one carefully curated AI briefing in your inbox—no endless feed, no filler.

















