Labs release new models and scientific agent benchmark leaderboards
The AI community launched the Terminal-Bench-Science 0.1 leaderboard for scientific research agents, alongside reports on upcoming large-scale model training at DeepSeek. Additionally, cost-optimized models like Gemini 3.8 Flash and new open-weight classifiers were deployed across serverless infrastructure.
10 independent accounts
70 posts
1 articles
3 labs
34,495 interactions
qwen3deepseek v4deepseekglmkv cachedeepseek v4 flashmixture of expertsqwen3.8
@AlexFinn
@ArtificialAnlys
@ClementDelangue
@EMostaque
@Hesamation
@OpenRouter
@PrismML
@alex_prompter
https://z.ai/ z.ai