CONSONANCE.for your information
Monday, 5 October 2026frenvi
CARD NO. 2026-09-24AI & AGENTS RADAR

Three voices, and it becomes news.

A radar over AI and agents. Every morning, and only what several independent sources are saying at once.

Being discussed

32 · topics

Subjects at least three independent accounts raised over the last seven days.

01 — multimodal · day 2 15 VOICES

OpenAI releases GPT-6 Sol and GPT-6 Luna alongside Tesseract video tools

OpenAI has introduced the GPT-6 Sol and GPT-6 Luna models, building on the capabilities of GPT-6 Astra to offer faster and more affordable options. Additionally, the company launched Tesseract, a video creative suite designed to give AI agents direct access to video editing tools.

15 independent accounts 80 posts 1 articles 8 labs 480,254 interactions
gpt-6
@AndrewCurran_ their topics on X ↗
@ArtificialAnlys their topics on X ↗
@Hesamation their topics on X ↗
@OpenAI their topics on X ↗
@OpenAIDevs their topics on X ↗
@OpenRouter their topics on X ↗
@PromptLLM their topics on X ↗
@aiedge_ their topics on X ↗
02 — evaluation · day 3 8 VOICES

SpaceXAI launches Grok 4.7 with high performance on the Artificial Analysis Intelligence Index

SpaceXAI has released Grok 4.7, which scores 46 on the Artificial Analysis Intelligence Index and places the lab in the top four AI labs. The model demonstrates improved coding agent performance while maintaining high speed and low cost.

8 independent accounts 116 posts 1 articles 1 labs 2,424,577 interactions
artificial analysisgrokcontext windowdeepsweopencodegrok voiceterminal-benchgrok imagine
@AlexFinn their topics on X ↗
@AndrewCurran_ their topics on X ↗
@ArtificialAnlys their topics on X ↗
@GithubProjects their topics on X ↗
@Hesamation their topics on X ↗
@aiedge_ their topics on X ↗
@dotey their topics on X ↗
@elonmusk their topics on X ↗
03 — architecture · day 2 10 VOICES

Anthropic releases Claude Opus 5.5 with 40% lower operating costs

Anthropic has introduced Claude Opus 5.5, the first model in the Claude 5.5 family, delivering performance comparable to Claude Fable 5.1 while cutting operating costs by 40% compared to Opus 5.

10 independent accounts 18 posts 2 labs 422,935 interactions
@AlexFinn their topics on X ↗
@AndrewCurran_ their topics on X ↗
@AnthropicAI their topics on X ↗
@Hesamation their topics on X ↗
@OpenRouter their topics on X ↗
@PromptLLM their topics on X ↗
@aiedge_ their topics on X ↗
@alex_prompter their topics on X ↗
04 — system_design NEW 9 VOICES

DigitalOcean announces public preview of Managed Agents

DigitalOcean has launched a public preview of Managed Agents, allowing users to run tools like Claude Code or Codex within runtime environments that pause automatically during idle periods.

9 independent accounts 60 posts 3 labs 180,825 interactions
codex
@CuiMao their topics on X ↗
@GithubProjects their topics on X ↗
@OpenAIDevs their topics on X ↗
@aiedge_ their topics on X ↗
@alex_prompter their topics on X ↗
@alliekmiller their topics on X ↗
@dani_avila7 their topics on X ↗
@dotey their topics on X ↗
05 — system_design · day 2 10 VOICES

NVIDIA provides accelerated computing infrastructure for Grok 4.7 and the AI ecosystem

NVIDIA's accelerated computing infrastructure is supporting models across the AI ecosystem, including powering SpaceXAI's Grok 4.7 for demanding coding and knowledge workflows.

10 independent accounts 48 posts 2 labs 190,994 interactions
nvidia
@ArtificialAnlys their topics on X ↗
@CuiMao their topics on X ↗
@NVIDIAAI their topics on X ↗
@SchmidhuberAI their topics on X ↗
@alex_prompter their topics on X ↗
@dair_ai their topics on X ↗
@elonmusk their topics on X ↗
@firstadopter their topics on X ↗
06 — agents NEW 8 VOICES

ChatGPT expands voice capabilities, multi-account plugins, and Chrome extensions

OpenAI updated ChatGPT voice capabilities to integrate plugins across email, calendar, and Slack. The platform also added support for linking multiple accounts with plugins and running Chrome extensions inside the desktop app.

8 independent accounts 66 posts 2 articles 4 labs 106,885 interactions
chatgpt
@AndrewBolis their topics on X ↗
@CuiMao their topics on X ↗
@OpenAI their topics on X ↗
@OpenAIDevs their topics on X ↗
@VibeMarketer_ their topics on X ↗
@alex_prompter their topics on X ↗
@dotey their topics on X ↗
@gdb their topics on X ↗
07 — multimodal NEW 10 VOICES

Google launches Gemini 3.8 Flash and Flash-Lite TTS audio model family

Google introduced Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS featuring over 2,000 production-ready voices and support for 100 languages. The release also extends real-time audio capabilities through Gemini 3.8 Live in search.

10 independent accounts 32 posts 2 articles 3 labs 51,900 interactions
gemini 3.8big bench audio
@Alibaba_Qwen their topics on X ↗
@AndrewCurran_ their topics on X ↗
@ArtificialAnlys their topics on X ↗
@EMostaque their topics on X ↗
@Google their topics on X ↗
@GoogleAIStudio their topics on X ↗
@OfficialLoganK their topics on X ↗
@aiedge_ their topics on X ↗
08 — architecture · day 2 6 VOICES

DeepSeek advances large model development and outlines hardware training roadmap

DeepSeek is currently training 2-trillion and 8-trillion parameter models, with plans to incorporate Huawei training chips in late 2026 or early 2027. The platform also integrated the DeepSeek V4.1 Flash model into developer tools.

6 independent accounts 28 posts 3 labs 121,304 interactions
dspydeepseek v4deepseek v4 flashdeepseek v4.1 flash
@EMostaque their topics on X ↗
@Hesamation their topics on X ↗
@elonmusk their topics on X ↗
@ollama their topics on X ↗
@op7418 their topics on X ↗
@teortaxesTex their topics on X ↗
@vllm_project their topics on X ↗
@TheAhmadOsman their topics on X ↗
09 — agents NEW 10 VOICES

Muse launches Mac app integrating versatile AI agent assistant for developers

The Muse app for Mac has been released, operating across files, calendar, notes, and messages. The platform also opened API connector access for developers and partnered with Shopify to streamline shopping workflows.

10 independent accounts 74 posts 1 articles 2 labs 215,272 interactions
muse sparkmeta ai
@AIatMeta their topics on X ↗
@VibeMarketer_ their topics on X ↗
@aiedge_ their topics on X ↗
@dani_avila7 their topics on X ↗
@emollick their topics on X ↗
@finkd their topics on X ↗
@firstadopter their topics on X ↗
@frankdegods their topics on X ↗
10 — system_design NEW 9 VOICES

Google introduces Googlebook laptop line powered by Android and ChromeOS stack

Google launched Googlebook, a new laptop featuring a 2.8K OLED touchscreen display and a 14-hour battery life. The device is built on the Android technology stack paired with ChromeOS desktop foundations.

9 independent accounts 73 posts 1 articles 2 labs 56,190 interactions
geminigemini 3.1 flash
@Alibaba_Qwen their topics on X ↗
@AndrewCurran_ their topics on X ↗
@ArtificialAnlys their topics on X ↗
@GeminiApp their topics on X ↗
@GithubProjects their topics on X ↗
@Google their topics on X ↗
@GoogleDeepMind their topics on X ↗
@Hesamation their topics on X ↗
11 — evaluation · day 5 11 VOICES

Anthropic partners with Accenture on independent evaluation and establishes biological lab

Anthropic partnered with Accenture for independent evaluations of frontier AI systems, committing at least $1 billion to building capacity. The company has also established a physical wet lab in San Francisco for biological research.

11 independent accounts 78 posts 1 articles 3 labs 104,057 interactions
anthropicclaude sonnetclaude haikuaime
@AlexFinn their topics on X ↗
@AndrewCurran_ their topics on X ↗
@AnthropicAI their topics on X ↗
@CuiMao their topics on X ↗
@Hesamation their topics on X ↗
@aiedge_ their topics on X ↗
@alex_prompter their topics on X ↗
@dair_ai their topics on X ↗
12 — multimodal NEW 6 VOICES

NVIDIA releases Nemotron 3 Diarization model for multi-speaker audio tracking

NVIDIA released the Nemotron 3 Diarization model with 100 million parameters, capable of tracking up to eight speakers and handling overlapping voices. The system identifies who spoke and when within multi-person audio transcripts.

6 independent accounts 43 posts 9 articles 4 labs 101,287 interactions
chain of thoughthugging face
@AndrewCurran_ their topics on X ↗
@ClementDelangue their topics on X ↗
@GithubProjects their topics on X ↗
@Gradio their topics on X ↗
@Hesamation their topics on X ↗
@NVIDIAAI their topics on X ↗
@_akhaliq their topics on X ↗
@alex_prompter their topics on X ↗
13 — agents · day 6 11 VOICES

Claude Code adds AGENTS.md configuration support and cloud session capabilities

Claude Code updated to version 2.1.277, adding support for AGENTS.md configuration files when CLAUDE.md is absent. Cloud sessions also exited research preview, enabling Claude to run parallel threads even after laptops are closed.

11 independent accounts 69 posts 1 articles 1 labs 226,932 interactions
claude code
@AlexFinn their topics on X ↗
@ArtificialAnlys their topics on X ↗
@ClaudeDevs their topics on X ↗
@CuiMao their topics on X ↗
@aiedge_ their topics on X ↗
@alex_prompter their topics on X ↗
@alliekmiller their topics on X ↗
@bcherny their topics on X ↗
14 — agents NEW 12 VOICES

Introduction of the CC agent and lobbying efforts over AI regulation

Google vừa công bố CC, một trợ lý AI được thiết kế để hỗ trợ các gia đình quản lý lịch trình và công việc chung. Trong một diễn biến khác, các lãnh đạo công nghệ lớn được cho là đã thúc ép chính quyền loại bỏ một đề xuất quản lý AI do người đứng đầu Google DeepMind đề xuất.

12 independent accounts 45 posts 1 articles 3 labs 109,756 interactions
googlegoogle deepminddemis hassabisjason wei
@Google their topics on X ↗
@Hesamation their topics on X ↗
@JeffDean their topics on X ↗
@OfficialLoganK their topics on X ↗
@aiedge_ their topics on X ↗
@alliekmiller their topics on X ↗
@boringmarketer their topics on X ↗
@dair_ai their topics on X ↗
15 — agents NEW 7 VOICES

Research on self-evolving ontologies and MCP integration for agents

Recent research demonstrates that incorporating a self-evolving ontology layer significantly improves agent performance on complex benchmarks. Meanwhile, the ecosystem continues to expand with new integrations and servers built around the Model Context Protocol.

7 independent accounts 29 posts 1 articles 3 labs 46,961 interactions
model context protocolmicrosoft research
@aiedge_ their topics on X ↗
@alex_prompter their topics on X ↗
@c_valenzuelab their topics on X ↗
@dair_ai their topics on X ↗
@dani_avila7 their topics on X ↗
@dotey their topics on X ↗
@ideabrowser their topics on X ↗
@op7418 their topics on X ↗
16 — multimodal · day 2 3 VOICES

PixVerse unveils the R2 world model and the impact of harness design

PixVerse has released R2, a real-time world model that allows users to interact with and edit dynamic environments using prompts. Additionally, recent studies highlight that employing an explicit graph harness can drastically improve a model's performance on long-horizon robotic tasks without altering the underlying weights.

3 independent accounts 11 posts 1 articles 3 labs 25,064 interactions
yann lecunworld model
@_akhaliq their topics on X ↗
@alex_prompter their topics on X ↗
@c_valenzuelab their topics on X ↗
@dair_ai their topics on X ↗
@tszzl their topics on X ↗
@10ayaGoldsen their topics on X ↗
@AmOptimistShow their topics on X ↗
@FuturLucide their topics on X ↗
17 — inference · day 6 8 VOICES

Release of Ternary Bonsai 2 27B and the Qwen3.8 real-time translation model

Ternary Bonsai 2 27B has been released, utilizing a 9x reduction in size compared to its Qwen3.8 base model while retaining 98.2% of its aggregate benchmark performance. Additionally, the Qwen3.8 ecosystem expanded with LiveTranslate, a simultaneous interpretation model built on an interleave architecture.

8 independent accounts 52 posts 2 articles 3 labs 120,411 interactions
kv cachemixture of expertsqwen3.8qwen3tokens per second
@Alibaba_Qwen their topics on X ↗
@ArtificialAnlys their topics on X ↗
@EMostaque their topics on X ↗
@OpenRouter their topics on X ↗
@PrismML their topics on X ↗
@godofprompt their topics on X ↗
@heyshrutimishra their topics on X ↗
@hwchase17 their topics on X ↗
18 — evaluation NEW 4 VOICES

MiMo-V2.6-Pro leads the open weights rankings on Artificial Analysis

MiMo-V2.6-Pro has debuted as the top open weights model on the Artificial Analysis Intelligence Index. Built on a classic grouped query attention architecture, it establishes a strong position on the intelligence versus cost per task Pareto frontier.

4 independent accounts 4 posts 1 labs 14,067 interactions
@ArtificialAnlys their topics on X ↗
@ClementDelangue their topics on X ↗
@rasbt their topics on X ↗
@teortaxesTex their topics on X ↗
19 — evaluation NEW 7 VOICES

Grok 4.7 xHigh competes at the top alongside updates on Kimi K3

Grok 4.7 xHigh has climbed close to the top spot on Artificial Analysis' AA-Briefcase benchmark, trailing the leader by a single point. Meanwhile, indicators point toward an upcoming Kimi K3.1 release, alongside demonstrations of running large models locally on standard CPUs.

7 independent accounts 34 posts 1 labs 19,598 interactions
glmkimi k3
@AndrewCurran_ their topics on X ↗
@Hesamation their topics on X ↗
@OpenRouter their topics on X ↗
@aiedge_ their topics on X ↗
@cline their topics on X ↗
@kimmonismus their topics on X ↗
@nutlope their topics on X ↗
@rohanpaul_ai their topics on X ↗
20 — agents NEW 3 VOICES

Discussions on automated AI research and multi-agent systems

The community discussed the potential of automating scientific research using multi-agent systems. Conversations centered on the prospects and implications of superintelligence driving autonomous breakthroughs in mathematics and engineering.

3 independent accounts 15 posts 1 labs 37,455 interactions
noam brown
@deanwball their topics on X ↗
@haider1 their topics on X ↗
@jerryjliu0 their topics on X ↗
@polynoamial their topics on X ↗
@scaling01 their topics on X ↗
@teortaxesTex their topics on X ↗
@tszzl their topics on X ↗
@dwarkesh_sp their topics on X ↗
21 — inference NEW 3 VOICES

Optimizing inference systems and accelerating execution for large language models

Zhipu reported that GLM-5.3-Flash runs across more than 10,000 domestic AI accelerators, achieving a 3.2x increase in throughput through automated optimization driven by an AI agent. Meanwhile, models like DiffusionGemma and Ternary-Bonsai received updates within vLLM and MLX frameworks.

3 independent accounts 23 posts 1 articles 2 labs 9,227 interactions
sglangvllm
@AndrewCurran_ their topics on X ↗
@AravSrinivas their topics on X ↗
@dotey their topics on X ↗
@teortaxesTex their topics on X ↗
@vllm_project their topics on X ↗
@PyTorch their topics on X ↗
@XiaomiMiMo their topics on X ↗
@aiDotEngineer their topics on X ↗
22 — architecture NEW 3 VOICES

DeepSeek trials large-scale server infrastructure and Xiaomi launches MiMo-V2.6

DeepSeek operates production-scale units spanning around 160 nodes, serving about 3 million sandboxes per day. Simultaneously, Xiaomi released the MiMo-V2.6 series featuring Pro and Flash models, with the Pro variant scoring 46 points on Artificial Analysis.

3 independent accounts 48 posts 2 labs 8,396 interactions
deepseekdeepseek r1moonshot ai
@Hesamation their topics on X ↗
@MatthewBerman their topics on X ↗
@dotey their topics on X ↗
@kimmonismus their topics on X ↗
@teortaxesTex their topics on X ↗
@54158zxc their topics on X ↗
@PeterDiamandis their topics on X ↗
@SemiAnalysis_ their topics on X ↗
23 — inference NEW 4 VOICES

Running local open-weight models on constrained hardware devices

Developers deployed a 27-billion-parameter ternary model at 1.72 bits per weight on a single graphics card with 12GB of VRAM. Additionally, updates brought llama.cpp's Metal kernels to transformers and added SGLang support for image generation models.

4 independent accounts 20 posts 3 labs 11,277 interactions
llama.cppllamaquantizationunslothgemma
@Alibaba_Qwen their topics on X ↗
@emollick their topics on X ↗
@teortaxesTex their topics on X ↗
@vllm_project their topics on X ↗
@ylecun their topics on X ↗
@DogukanUrker their topics on X ↗
@HuggingModels their topics on X ↗
@NewsFromGoogle their topics on X ↗
24 — system_design NEW 3 VOICES

Alibaba announces large-scale AI infrastructure plans and RecreationWorld sandbox

Alibaba targeted 20 gigawatts of global data center capacity by 2032 alongside plans for a 5 to 10 trillion parameter Qwen model series. The Qwen team also released RecreationWorld, a five-platform sandbox environment for computer-use agents.

3 independent accounts 15 posts 1 labs 1,804 interactions
bytedancealibabaalibaba qwen
@AravSrinivas their topics on X ↗
@Hesamation their topics on X ↗
@alex_prompter their topics on X ↗
@rohanpaul_ai their topics on X ↗
@teortaxesTex their topics on X ↗
@AlibabaGroup their topics on X ↗
@HuggingPapers their topics on X ↗
@perplexity_ai their topics on X ↗
25 — agents NEW 3 VOICES

Improving skill optimization methods for AI agents

Researchers introduced a method allowing AI agents to improve skill prompts by ranking candidates, achieving 40% to 70% lower token costs compared to conventional methods. NVIDIA also detailed an approach to compile public agent skills into reinforcement learning environments.

3 independent accounts 9 posts 1,980 interactions
agent skills
@NVIDIAAI their topics on X ↗
@aiedge_ their topics on X ↗
@dair_ai their topics on X ↗
@dani_avila7 their topics on X ↗
@claudeskills101 their topics on X ↗
26 — evaluation NEW 3 VOICES

Evaluating safety refusal rates on coding agent benchmarks

Artificial Analysis introduced safety refusal reporting in version 1.5 of its Coding Agent Index to help explain model behavioral differences. Data showed distinct variations in safety refusal rates among evaluated models during coding tasks.

3 independent accounts 5 posts 1,583 interactions
devin
@ArtificialAnlys their topics on X ↗
@Hesamation their topics on X ↗
@NVIDIAAI their topics on X ↗
@0x15f their topics on X ↗
27 — agents NEW 3 VOICES

Automated learning applications and improved retrieval methods for finance

A joint university study introduced a path for financial agents to learn from SEC filing errors by requiring all new behaviors to pass regression tests first. New adaptive retrieval frameworks were also released to better handle specific data queries.

3 independent accounts 9 posts 1,393 interactions
retrieval augmented generation
@AndrewBolis their topics on X ↗
@GithubProjects their topics on X ↗
@alex_prompter their topics on X ↗
@rohanpaul_ai their topics on X ↗
@boltzbit their topics on X ↗

Worth reading closely

24 reads

Papers and writeups, read from their abstracts, ranked by relevance to someone building agents and backends.

01 — agents 217 upvotes

Realtime-Venus: A full-duplex interaction system with asynchronous delegation

QUESTION — How can a full-duplex interaction system be built to integrate continuous perception, conversational control, and asynchronous background task execution using 9B models?

Realtime-Venus-Omni achieves the highest scores on six of eight video benchmarks, including StreamingBench (70.2%), OVO-Bench (64.7%), and Daily-Omni (81.3%).

AdinaY · 12 Sept 2026 read the original ↗
02 — system_design 67 upvotes

Document Retrieval-Aware Chunking (D-RAC): Universal Retrieval-Aware Ingestion of Enterprise Documents via PDF Normalization and Multimodal Markdown Conversion

QUESTION — How can heterogeneous enterprise document formats (PDFs, Word, scans) be converted and chunked into retrieval-optimized Markdown while minimizing token costs and processing time?

On the 236-document, 795-page PDF subset of the RAG-Multi-Corpus benchmark, D-RAC converts and chunks the entire corpus in 72 minutes with zero errors, producing 1,748 retrieval-ready chunks.

udayallu · 21 Sept 2026 read the original ↗
06 — training 23 upvotes

PACT: From Credit Assignment to Critic Alignment

QUESTION — How can token-level credit be mathematically defined and leveraged to improve actor-critic training in LLM post-training?

In agentic mathematical reasoning, PACT achieves 72.87% average accuracy across four benchmarks, outperforming GRPO and PPO by 8.80 and 13.16 percentage points, respectively.

Eclipse2001 · 22 Sept 2026 read the original ↗
10 — inference 42 upvotes

JEV-as-a-Judge: Accept When Confident, Escalate When Unsure

QUESTION — How can a decision-only judge reduce the inference cost of LLM-as-a-judge while preserving evaluation accuracy?

JEV is within three percentage points of a state-of-the-art LLM judge on ordinary preference and evidence-grounded factuality at 0.36% of the comparator's fee.

yubol · 22 Sept 2026 read the original ↗
11 — training 17 upvotes

Towards Full Pipeline FP8 Reinforcement Learning for LLMs

QUESTION — What causes training instability in full-pipeline FP8 reinforcement learning for LLMs, and how can it be resolved?

Compounded FP8 quantization noise distorts the importance ratio, pushing negative-advantage tokens outside the trust region and zeroing out their gradients.

CFC · 19 Sept 2026 read the original ↗
15 — training 13 upvotes

From Pretraining to Proficiency: Real-World Subtask RL for Long-Horizon Manipulation with Minimal Human Intervention

QUESTION — How can reinforcement learning fine-tuning be effectively applied to long-horizon real-world manipulation tasks with minimal human intervention?

On bimanual YAM and single-arm Franka tasks, PARTS improves complete-task success from 32% to 61% and from 50% to 95%, respectively, using tens of minutes of real-world RL rollouts per task on average.

Sichang0621 · 18 Sept 2026 read the original ↗
16 — training 101 upvotes

RULER: Instance-aware Rubric Rewards for SVG Generation

QUESTION — How can reinforcement learning policies be optimized for open-ended SVG code generation without relying on ground truth data or human preference labels?

Prompting a vision-language judge with a multi-axis rubric correlates with human judgments far better than scalar metrics.

hangyuran · 21 Sept 2026 read the original ↗
17 — architecture 49 upvotes

The Past Frames the Future: Memory for Autoregressive Video Generation

QUESTION — What memory mechanisms preserve temporal persistence and historical information in autoregressive video generation under bounded context limits?

This paper presents a systematic and comprehensive review of memory mechanisms in autoregressive (AR) video generation. As generated sequences expand, models face strict constraints on context windows, storage, and compute, causing critical historical information to leave the active context. The authors formulate memory operationally as persistent historical information maintained across outer AR steps to influence future generation. The literature is organized through five complementary perspectives: Forms, Functions, Operations, Learning, and Evaluation, establishing a structured foundation for building reliable memory-conditioned video generation systems.

Harold328 · 23 Sept 2026 read the original ↗
23 — evaluation 45 upvotes

HappyWorld-Bench

QUESTION — How can we comprehensively evaluate the consistency and responsiveness of world models as agents interact and explore?

Spatial models achieve at best 70.14% placement accuracy and 73.33% edit execution.

CheeryLJH · 21 Sept 2026 read the original ↗

Hands-on

2026-09-24

Repositories climbing on GitHub today, scored on the day's momentum, rank, adoption, freshness and project health.

01 anthropics/financial-services +664 A collection of LLM integration examples for backend engineers building secure financial applications with Claude. Python
★ 37,102
02 davila7/claude-code-templates +389 A CLI utility for configuring and monitoring Claude Code, designed for developers optimizing their terminal-based AI coding workflows. Python
★ 31,625
03 google/ax +1,543 Google's Go-based agentic orchestration runtime for backend engineers building robust, large-scale multi-agent enterprise systems. Go
★ 9,465
04 obra/superpowers +474 A development methodology and skills framework designed to orchestrate software-engineering AI agents. Shell
★ 290,870
05 dream-num/univer +1,142 An integrated spreadsheet, doc, and canvas runtime for engineers building agents that need programmatic Office document manipulation. TypeScript
★ 16,544
06 pbakaus/impeccable +304 A design system specification for developers to guide their coding agents in generating production-grade, aesthetically pleasing user interfaces. JavaScript
★ 70,514
07 agent-substrate/substrate +558 A Go-based core systems infrastructure for backend engineers building scalable execution environments for autonomous AI agents. Go
★ 3,618
08 superdesigndev/treg +506 An OpenRouter-style proxy specifically for agent tool calls, helping backend engineers route and manage LLM tool executions. Python
★ 2,861
09 browser-use/video-use +746 A coding agent framework for automated video editing, built for engineers developing LLM-driven multimedia processing pipelines. Python
★ 26,647
10 BuilderIO/agent-native +87 A TypeScript framework that helps backend engineers design and deploy complex agentic applications. TypeScript
★ 6,678
11 DeusData/codebase-memory-mcp +190 High-performance code intelligence MCP server. Indexes codebases into a persistent knowledge graph — average repo in milliseconds. 158 languages, sub-ms queries, 99% fewer tokens. Single static binary, zero dependencies. C
★ 44,714
12 strands-agents/harness-sdk +115 An open-source SDK for building and orchestrating production-grade agent harnesses in Python and TypeScript. Python
★ 7,982
13 HKUDS/CLI-Anything +57 A framework that wraps standard software into agent-native CLI interfaces for seamless automation. Python
★ 50,065
14 TNT-Likely/PanWatch +95 盯盘侠 PanWatch · 自托管 AI 盯盘助手,集成 TradingAgents 多 Agent 投资决策 | A股/港股/美股实时监控、持仓管理、智能分析、全渠道推送 Python
★ 1,666
15 harry7557558/spirula-studio +69 Cross-vendor 3D Gaussian Splatting trainer - video to splat to mesh, Vulkan or CUDA. C++
★ 824

Claims

1 claims

Checkable assertions with their sources and their contradictions. Each stands on at least two independent sources, or a person read it first; the stamp says which.

TWO SOURCES · NOT REREAD
01 · 24 Sept 2026 · claude · 2 sources

Claude discovered a previously uncharacterized enzyme system with features reminiscent of CRISPR after AI agents searched through a DNA sequence database.

Condition: With 950 agents running over 21 hours.

evidence · @nc_frey
“We’ve set up a molecular biology lab at Anthropic and we’re announcing our first discovery! Claude discovered a new CRISPR-like enzyme. 950 agents spent 21 hours searching through a database of DNA sequences until one of the agents found something striking”
evidence · @BoWang87
“About 950 Claude agent sessions ran over 21 hours, surveyed ~200,000 reverse transcriptases, and produced 19 reports.”
Still to check
  • Re-examine the surveyed DNA database and experimentally verify the enzyme's activity in the lab.
↑