Today’s lead · x.com
Anthropic reportedly tightens Claude access against Chinese-company workarounds
Cointelegraph, citing the Financial Times, says Anthropic is moving to close loopholes that let Chinese companies access Claude through workarounds.

Top signals
5 moreTools & repos
5 selected
Antigravity CLI
Terminal-native AI coding agent that brings Antigravity 2.0-style multi-step reasoning, multi-file editing, tool calling, command execution, persistent history, SSH-friendly auth, and session export into a TUI. Useful for keyboard-first and remote workflows, but it should be governed like any code-executing agent.
openai/codex-plugin-cc
OpenAI's Claude Code plugin lets users call Codex from inside Claude Code to review code or delegate tasks, a useful bridge for teams comparing or combining coding agents in one workflow.

Osloq
Osloq is an AI agent for reproducing GitHub issues, aimed at turning vague bug reports into repeatable cases before maintainers spend engineering time debugging them.
Taste Skill v2
Open-source frontend skill files for Cursor, Claude Code, Codex, Gemini CLI, v0, Lovable, OpenCode, and other agents. The v2 ruleset reads the brief, chooses an appropriate design direction or design system, audits existing UI, and bans generic AI-generated frontend patterns.

Glaze by Raycast
Raycast's Glaze lets users create small Mac apps by chatting with AI, a fit for lightweight internal tools and personal automations where a full product build would be overkill.
Blogs worth your time
4 reads
Qwen3-Omni serving in vLLM-Omni
The vLLM team breaks Qwen3-Omni serving into a Thinker, Talker, and Code2Wav pipeline, then shows how batching, CUDA graphs, async handoffs, and speech-stage replicas reduce first-audio latency from about 6s to about 0.6s and lift throughput roughly 5.4x under load.

Modern GPU Programming For MLSys
Tianqi Chen's open book builds a practical path from GPU hardware mental models to Blackwell-era kernels, using GEMM and FlashAttention examples to teach layouts, TMA, tensor cores, barriers, persistent scheduling, and warp specialization.

Open sharded inference of a 229B MoE over the public internet
c0mpute's technical report runs MiniMax-M2.5 across five consumer RTX 5090s in five European countries, measuring 12.6 tok/s interactive and 194 tok/s batched with cryptographic receipts on every request. The useful takeaway is where decentralized inference breaks: host CPUs and speculative decoding over high-latency links, not only raw GPU capacity.

Google DeepMind and A24 announce first-of-its-kind research partnership
Google DeepMind and A24 announced a multi-project research collaboration, with Google also investing in A24, to test AI-assisted creative workflows directly with filmmakers. There are no technical outputs yet, but it is a useful example of frontier labs embedding inside vertical workflows instead of shipping generic tools in isolation.
Funding & acquisitions
1 movesBell, Cohere, Hypertec and BUZZ HPC
The companies announced a Canadian sovereign AI infrastructure effort combining Bell's national platform, Cohere's enterprise AI, Hypertec's Canadian-built GPU servers, and BUZZ HPC's AI factory and sovereign AI cloud.
Bengaluru radar
1 eventsBangalore Tech Mixer and Social (Tech / AI / Data / IT)
After-work networking for Bangalore software, data, AI, cybersecurity, product, and startup people at We:Neighborhood. Useful if you want local hiring, referral, and peer-network access rather than a formal talk.




