OpenAI System CardAddendum to GPT-6 Astra System Card: GPT-6.1 Sol
2026-09-29 · 48 页
Anthropic System CardSystem Card: Claude Sonnet 5.5
2026-09-28 · 148 页
Anthropic System CardClaude Opus 5.5 System Card
2026-09-22 · 230 页
xAI Model CardGrok 4.7 Model Card
2026-09-21 · 30 页
Tencent / Hunyuan Technical ReportWeVisDoc: From Coverage to Capability for Robust End-to-End Document Parsing
2026-09-17 · arxiv-v1 · 35 页
DeepSeek Technical ReportDeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression
2026-09-17 · arxiv-v1 · 51 页
StepFun Technical ReportStepAudio 3 Realtime Technical Report
2026-09-12 · arxiv-v2 · 27 页
Cohere Technical ReportNorth Small Translate: Advanced Cost-Effective Translation (Cohere CAT+)
2026-09-12 · arxiv-v1 · 14 页
StepFun Technical ReportStepAudio 3 Music Technical Report
2026-09-11 · arxiv-v1 · 18 页
StepFun Technical ReportStepAudio 3 Gen Technical Report
2026-09-11 · arxiv-v1 · 21 页
OpenAI System CardChatGPT Images 2.5 System Card
2026-09-08 · 8 页
Alibaba / Qwen / Wan Technical ReportQwen-Audio-3.0-ASR Technical Report
2026-09-07 · arxiv-v2 · 21 页
Google / DeepMind Technical ReportWeatherNext 3: Increasing Resolution and Performance of Global Weather Models with Raw Observations
2026-09-03 · arxiv-v1 · 43 页
OpenAI System CardGPT-6 Astra System Card
2026-09-03 · 175 页
NVIDIA Technical ReportPost-Training Language Models for Gold-Medal Performance in Coding Competitions
2026-09-02 · arxiv-v2 · 21 页
Anthropic System CardClaude Fable 5.1 and Mythos 5.1 System Card
2026-09-01 · 212 页
Alibaba / Qwen / Wan Technical ReportQwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving
2026-08-31 · arxiv-v1 · 40 页
Alibaba / Qwen / Wan Technical ReportOn the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability
2026-08-26 · 28 页
xAI Model CardGrok 4.6 Model Card
2026-08-12 · 42 页
Alibaba / Qwen / Wan Technical ReportWan-Animate-2: Pushing the Application Boundaries of Character Animation
2026-08-06 · arxiv-v2 · 14 页
OpenAI System CardGPT-5.6 — August Updates
2026-08-06 · 30 页
Tencent / Hunyuan Technical ReportWorldClaw: Agentic 3D Open-World Generation at Scale
2026-08-05 · arxiv-v1 · 38 页
Tencent / Hunyuan Technical ReportHunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing
2026-08-03 · arxiv-v3 · 33 页
Alibaba / Qwen / Wan Technical ReportQwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents
2026-07-30 · arxiv-v1 · 56 页
Mistral AI Technical ReportShieldstral
2026-07-28 · arxiv-v2 · 24 页
Anthropic System CardClaude Opus 5 System Card
2026-07-24 · 198 页
Mistral AI Technical ReportRobostral Navigate
2026-07-22 · arxiv-v3 · 12 页
xAI Model CardGrok 4.5 Model Card
2026-07-14 · 29 页
Alibaba / Qwen / Wan Technical ReportWan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation
2026-07-10 · arxiv-v3 · 17 页
OpenAI System CardGPT-5.6 System Card
2026-07-09 · 82 页
OpenAI System CardGPT-Live System Card
2026-07-08 · 7 页
NVIDIA Technical ReportUnified Audio Intelligence Without Regressing on Text Intelligence
2026-07-06 · arxiv-v2 · 41 页
Tencent / Hunyuan Technical ReportHunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
2026-07-06 · arxiv-v3 · 41 页
NVIDIA Technical ReportNemotron-Labs-3-Puzzle-75B-A9B: Compressing Hybrid MoE LLMs
2026-07-05 · arxiv-v2 · 25 页
Mistral AI Technical ReportLeanstral
2026-07-02 · 12 页
Google / DeepMind Technical ReportGemma 4 Technical Report
2026-07-02 · arxiv-v2 · 17 页
Anthropic System CardClaude Sonnet 5 System Card
2026-06-30 · 146 页
OpenAI System CardGPT-5.6 Preview System Card
2026-06-26 · 77 页
Alibaba / Qwen / Wan Technical ReportQwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System
2026-06-16 · arxiv-v3 · 37 页
Alibaba / Qwen / Wan Technical ReportQwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models
2026-06-16 · arxiv-v2 · 44 页
MiniMax Technical ReportMiniMax Sparse Attention
2026-06-11 · arxiv-v2 · 30 页
Anthropic System CardClaude Fable 5 and Mythos 5 System Card
2026-06-09 · 317 页
OpenAI System CardGPT-Rosalind-5.5 System Card
2026-06-03 · 11 页
Microsoft / Phi / MAI Technical ReportMAI-Thinking-1: Building a Hill-Climbing Machine
2026-06-02 · 109 页
NVIDIA Technical ReportCosmos 3: Omnimodal World Models for Physical AI
2026-06-01 · 139 页
Alibaba / Qwen / Wan Technical ReportQwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
2026-05-28 · arxiv-v2 · 34 页
Anthropic System CardClaude Opus 4.8 System Card
2026-05-28 · 246 页
MiniMax Technical ReportThe MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence
2026-05-26 · arxiv-v2 · 35 页
Google / DeepMind Technical ReportGemini Embedding 2: A Native Multimodal Embedding Model from Gemini
2026-05-26 · arxiv-v1 · 21 页
StepFun Technical ReportStepAudio 2.5 Technical Report
2026-05-22 · arxiv-v1 · 19 页
Tencent / Hunyuan Technical ReportHy-MT2: A Family of Fast, Efficient and Powerful Multilingual Translation Models in the Wild
2026-05-21 · arxiv-v2 · 15 页
Meta System CardMuse Spark Safety & Preparedness Report
2026-05-14 · arxiv-v1 · 160 页
Alibaba / Qwen / Wan Technical ReportQwen-Image-2.0 Technical Report
2026-05-11 · arxiv-v1 · 30 页
OpenAI System CardGPT-5.5 Instant System Card
2026-05-05 · 21 页
StepFun Technical ReportStep-Audio-R1.5 Technical Report
2026-04-28 · arxiv-v2 · 9 页
NVIDIA Technical ReportNemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence
2026-04-27 · arxiv-v2 · 27 页
DeepSeek Model CardDeepSeek V4 Technical Documentation (Model Card)
2026-04-27 · 8 页
DeepSeek Technical ReportDeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
2026-04-26 · arxiv-v1 · 58 页
OpenAI System CardGPT-5.5 System Card
2026-04-23 · 46 页
ByteDance / Seed Technical ReportSeed3D 2.0: Advancing High-Fidelity Simulation-Ready 3D Content Generation
2026-04-22 · arxiv-v1 · 18 页
OpenAI System CardChatGPT Images 2.0 System Card
2026-04-21 · 5 页
Alibaba / Qwen / Wan Technical ReportQwen3.5-Omni Technical Report
2026-04-17 · arxiv-v2 · 28 页
DeepSeek Model CardDeepSeek V3.2 Technical Documentation (Model Card)
2026-04-17 · 7 页
Anthropic System CardClaude Opus 4.7 System Card
2026-04-16 · 232 页
ByteDance / Seed Technical ReportSeedance 2.0: Advancing Video Generation for World Complexity
2026-04-15 · arxiv-v1 · 26 页
Anthropic System CardMythos Preview System Card
2026-04-07 · 245 页
xAI System CardGrok 4.20 System Card
2026-04-07 · 8 页
Google / DeepMind Technical ReportMedGemma 1.5 Technical Report
2026-04-06 · arxiv-v2 · 23 页
Mistral AI Technical ReportVoxtral TTS
2026-03-26 · arxiv-v2 · 16 页
NVIDIA Technical ReportNemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
2026-03-19 · arxiv-v2 · 63 页
Meta Technical ReportV-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning
2026-03-15 · arxiv-v3 · 37 页
Cohere Technical ReportTiny Aya: Bridging Scale and Multilingual Depth
2026-03-12 · arxiv-v1 · 50 页
智谱 / Z.ai Technical ReportGLM-OCR Technical Report
2026-03-11 · arxiv-v2 · 17 页
OpenAI System CardGPT-5.4 Thinking System Card
2026-03-05 · 38 页
Microsoft / Phi / MAI Technical ReportPhi-4-reasoning-vision-15B Technical Report
2026-03-04 · arxiv-v1 · 21 页
OpenAI System CardGPT-5.3 Instant System Card
2026-03-02 · 5 页
Alibaba / Qwen / Wan Technical ReportQwen3-Coder-Next Technical Report
2026-02-28 · arxiv-v1 · 23 页
Meta Technical ReportSAM 3D Body: Robust Full-Body Human Mesh Recovery
2026-02-17 · arxiv-v1 · 20 页
智谱 / Z.ai Technical ReportGLM-5: from Vibe Coding to Agentic Engineering
2026-02-17 · arxiv-v2 · 40 页
Anthropic System CardClaude Sonnet 4.6 System Card
2026-02-17 · 135 页
Mistral AI Technical ReportVoxtral Realtime
2026-02-11 · arxiv-v3 · 18 页
StepFun Technical ReportStep 3.5 Flash: Open Frontier-Level Intelligence with 11B Active Parameters
2026-02-11 · arxiv-v2 · 67 页
Anthropic System CardClaude Opus 4.6 System Card
2026-02-06 · 213 页
OpenAI System CardGPT-5.3-Codex System Card
2026-02-05 · 31 页
Moonshot AI / Kimi Technical ReportKimi K2.5 Technical Report
2026-02-02 · 31 页
Alibaba / Qwen / Wan Technical ReportQwen3-ASR Technical Report
2026-01-29 · arxiv-v2 · 16 页
DeepSeek Technical ReportDeepSeek-OCR 2: Visual Causal Flow
2026-01-28 · arxiv-v1 · 15 页
Amazon / Nova System CardEvaluating Nova 2.0 Lite model under Amazon's Frontier Model Safety Framework
2026-01-27 · arxiv-v1 · 9 页
Alibaba / Qwen / Wan Technical ReportQwen3-TTS Technical Report
2026-01-22 · arxiv-v1 · 14 页
StepFun Technical ReportSTEP3-VL-10B Technical Report
2026-01-14 · arxiv-v2 · 50 页
Google / DeepMind Technical ReportTranslateGemma Technical Report
2026-01-13 · arxiv-v3 · 12 页
Mistral AI Technical ReportMinistral 3
2026-01-13 · arxiv-v1 · 14 页
Alibaba / Qwen / Wan Technical ReportQwen3-VL-Embedding and Qwen3-VL-Reranker Technical Report
2026-01-08 · 23 页
StepFun Technical ReportStep-DeepResearch Technical Report
2025-12-23 · arxiv-v4 · 27 页
Meta Technical ReportPushing the Frontier of Audiovisual Perception with Large-Scale Multimodal Correspondence Learning
2025-12-22 · arxiv-v1 · 38 页
Meta Technical ReportSAM Audio: Segment Anything in Audio
2025-12-19 · arxiv-v1 · 57 页
OpenAI System CardAddendum to GPT-5.2 System Card: GPT-5.2-Codex
2025-12-18 · 22 页
ByteDance / Seed Model CardSeed1.8 Model Card: Towards Generalized Real-World Agency
2025-12-17 · 48 页
Alibaba / Qwen / Wan Technical ReportQwen-Image-Layered: Towards Inherent Editability via Layer Decomposition
2025-12-17 · arxiv-v1 · 12 页
Google / DeepMind Technical ReportT5Gemma 2: Seeing, Reading, and Understanding Longer
2025-12-16 · arxiv-v2 · 13 页
智谱 / Z.ai Technical ReportGLM-TTS Technical Report
2025-12-16 · arxiv-v1 · 14 页
MiniMax Technical ReportTowards Scalable Pre-training of Visual Tokenizers for Generation
2025-12-15 · arxiv-v2 · 16 页
ByteDance / Seed Model CardSeedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model
2025-12-15 · arxiv-v3 · 11 页
NVIDIA Technical ReportNemotron-Cascade: Scaling Cascaded Reinforcement Learning for General-Purpose Reasoning Models
2025-12-15 · arxiv-v2 · 60 页
NVIDIA Technical ReportNVIDIA Nemotron 3: Efficient and Open Intelligence
2025-12-15 · 13 页
NVIDIA Technical ReportNVIDIA Nemotron 3 Nano Technical Report
2025-12-15 · 40 页
OpenAI System CardUpdate to GPT-5 System Card: GPT-5.2
2025-12-11 · 27 页
Alibaba / Qwen / Wan Technical ReportWan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance
2025-12-09 · arxiv-v1 · 22 页
DeepSeek Technical ReportDeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
2025-12-02 · arxiv-v1 · 23 页
Amazon / Nova Technical ReportAmazon Nova 2: Multimodal Reasoning and Generation Models
2025-12-02 · 26 页
StepFun Technical ReportReasonEdit: Towards Reasoning-Enhanced Image Editing Models
2025-11-27 · arxiv-v2 · 18 页
DeepSeek Technical ReportDeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
2025-11-27 · 19 页
Alibaba / Qwen / Wan Technical ReportQwen3-VL Technical Report
2025-11-26 · arxiv-v2 · 42 页
Tencent / Hunyuan Technical ReportHunyuanVideo 1.5 Technical Report
2025-11-24 · arxiv-v2 · 14 页
Tencent / Hunyuan Technical ReportHunyuanOCR Technical Report
2025-11-24 · arxiv-v2 · 36 页
Anthropic System CardClaude Opus 4.5 System Card
2025-11-24 · 153 页
Meta Technical ReportSAM 3D: 3Dfy Anything in Images
2025-11-20 · arxiv-v2 · 44 页
Meta Technical ReportSAM 3: Segment Anything with Concepts
2025-11-20 · arxiv-v2 · 78 页
StepFun Technical ReportStep-Audio-R1 Technical Report
2025-11-19 · arxiv-v2 · 22 页
OpenAI System CardGPT-5.1-Codex-Max System Card
2025-11-18 · 27 页
xAI Model CardGrok 4.1 Model Card
2025-11-17 · 6 页
OpenAI System CardGPT-5.1 Instant and GPT-5.1 Thinking System Card Addendum
2025-11-12 · 5 页
StepFun Technical ReportStep-Audio-EditX Technical Report
2025-11-05 · arxiv-v2 · 13 页
Cohere Technical ReportCommand-A-Translate: Raising the Bar of Machine Translation with Difficulty Filtering
2025-11 · 11 页
Moonshot AI / Kimi Technical ReportKimi Linear: An Expressive, Efficient Attention Architecture
2025-10-30 · arxiv-v2 · 28 页
NVIDIA Technical ReportAlpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail
2025-10-30 · arxiv-v2 · 42 页
OpenAI Technical ReportTechnical Report: Performance and baseline evaluations of gpt-oss-safeguard-120b and gpt-oss-safeguard-20b
2025-10-29 · 10 页
Alibaba / Qwen / Wan Technical ReportTongyi DeepResearch Technical Report
2025-10-28 · arxiv-v3 · 23 页
Amazon / Nova Technical ReportAmazon Nova Multimodal Embeddings: Technical Report and Model Card
2025-10-28 · 5 页
OpenAI System CardAddendum to GPT-5 System Card: Sensitive Conversations
2025-10-27 · 4 页
DeepSeek Technical ReportDeepSeek-OCR: Contexts Optical Compression
2025-10-21 · arxiv-v1 · 22 页
Alibaba / Qwen / Wan Technical ReportQwen3Guard Technical Report
2025-10-16 · arxiv-v1 · 28 页
Anthropic System CardClaude Haiku 4.5 System Card
2025-10 · 39 页
OpenAI System CardSora 2 System Card
2025-09-30 · 7 页
Tencent / Hunyuan Technical ReportHunyuanImage 3.0 Technical Report
2025-09-28 · arxiv-v3 · 17 页
Moonshot AI / Kimi Technical ReportKimi-Dev: Agentless Training as Skill Prior for SWE-Agents
2025-09-27 · arxiv-v3 · 68 页
ByteDance / Seed Technical ReportSeedream 4.0: Toward Next-generation Multimodal Image Generation
2025-09-24 · arxiv-v3 · 19 页
Google / DeepMind Technical ReportEmbeddingGemma: Powerful and Lightweight Text Representations
2025-09-24 · arxiv-v3 · 18 页
Alibaba / Qwen / Wan Technical ReportQwen3-Omni Technical Report
2025-09-22 · arxiv-v1 · 25 页
Alibaba / Qwen / Wan Technical ReportWan-Animate: Unified Character Animation and Replacement with Holistic Replication
2025-09-17 · arxiv-v1 · 17 页
OpenAI System CardAddendum to GPT-5 system card: GPT-5-Codex
2025-09-15 · 7 页
Tencent / Hunyuan Technical ReportHunyuan-MT Technical Report
2025-09-05 · arxiv-v2 · 18 页
Anthropic System CardClaude Sonnet 4.5 System Card
2025-09 · 149 页
Alibaba / Qwen / Wan Technical ReportWan-S2V: Audio-Driven Cinematic Video Generation
2025-08-26 · arxiv-v1 · 11 页
NVIDIA Technical ReportJet-Nemotron: Efficient Language Model with Post Neural Architecture Search
2025-08-21 · arxiv-v3 · 20 页
NVIDIA Technical ReportNVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model
2025-08-20 · arxiv-v4 · 43 页
StepFun Technical ReportNextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale
2025-08-14 · arxiv-v2 · 25 页
Meta Technical ReportDINOv3
2025-08-13 · arxiv-v1 · 67 页
OpenAI Model Cardgpt-oss-120b & gpt-oss-20b Model Card
2025-08-08 · arxiv-v1 · 35 页
智谱 / Z.ai Technical ReportGLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models
2025-08-08 · arxiv-v1 · 26 页
OpenAI System CardGPT-5 System Card
2025-08-07 · 60 页
ByteDance / Seed Technical ReportSeed Diffusion: A Large-Scale Diffusion Language Model with High-Speed Inference
2025-08-04 · arxiv-v1 · 11 页
Alibaba / Qwen / Wan Technical ReportQwen-Image Technical Report
2025-08-04 · arxiv-v1 · 46 页
Anthropic System CardClaude Opus 4.1 System Card
2025-08 · 23 页
Moonshot AI / Kimi Technical ReportKimi K2: Open Agentic Intelligence
2025-07-28 · arxiv-v2 · 32 页
StepFun Technical ReportStep-3 is Large yet Affordable: Model-system Co-design for Cost-effective Decoding
2025-07-25 · arxiv-v1 · 18 页
StepFun Technical ReportStep-Audio 2 Technical Report
2025-07-22 · arxiv-v3 · 21 页
Mistral AI Technical ReportVoxtral
2025-07-17 · arxiv-v1 · 17 页
OpenAI System CardChatGPT Agent System Card
2025-07-17 · 42 页
Microsoft / Phi / MAI Technical ReportDecoder-Hybrid-Decoder Architecture for Efficient Reasoning with Long Generation
2025-07-09 · arxiv-v3 · 35 页
Google / DeepMind Technical ReportMedGemma Technical Report
2025-07-07 · arxiv-v4 · 60 页
Google / DeepMind Technical ReportGemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025-07-07 · arxiv-v6 · 73 页
Amazon / Nova System CardEvaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
2025-07-07 · arxiv-v1 · 10 页
智谱 / Z.ai Technical ReportGLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
2025-07-01 · arxiv-v6 · 42 页
Tencent / Hunyuan Technical ReportHunyuan-A13B Technical Report
2025-06-27 · 14 页
Tencent / Hunyuan Technical ReportHunyuan3D 2.1: From Images to High-Fidelity 3D Assets with Production-Ready PBR Material
2025-06-18 · arxiv-v1 · 14 页
MiniMax Technical ReportMiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
2025-06-16 · arxiv-v1 · 22 页
NVIDIA Technical ReportAceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy
2025-06-16 · arxiv-v1 · 23 页
Mistral AI Technical ReportMagistral
2025-06-12 · arxiv-v1 · 23 页
Meta Technical ReportV-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
2025-06-11 · arxiv-v1 · 48 页
StepFun Technical ReportStep-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language Model
2025-06-10 · arxiv-v2 · 12 页
ByteDance / Seed Technical ReportSeedance 1.0: Exploring the Boundaries of Video Generation Models
2025-06-10 · arxiv-v2 · 26 页
Alibaba / Qwen / Wan Technical ReportQwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
2025-06-05 · arxiv-v3 · 14 页
NVIDIA Technical ReportAceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
2025-05-22 · arxiv-v3 · 23 页
OpenAI System CardAddendum to OpenAI o3 and o4-mini system card: Codex
2025-05-16 · 8 页
Alibaba / Qwen / Wan Technical ReportWorldPM: Scaling Human Preference Modeling
2025-05-15 · arxiv-v2 · 34 页
Alibaba / Qwen / Wan Technical ReportQwen3 Technical Report
2025-05-14 · arxiv-v1 · 35 页
Cohere Technical ReportAya Vision: Advancing the Frontier of Multilingual Multimodality
2025-05-13 · arxiv-v1 · 76 页
StepFun Technical ReportStep1X-3D: Towards High-Fidelity and Controllable Generation of Textured 3D Assets
2025-05-12 · arxiv-v1 · 23 页
MiniMax Technical ReportMiniMax-Speech: Intrinsic Zero-Shot Text-to-Speech with a Learnable Speaker Encoder
2025-05-12 · arxiv-v1 · 20 页
ByteDance / Seed Technical ReportSeed1.5-VL Technical Report
2025-05-11 · arxiv-v1 · 77 页
NVIDIA Technical ReportLlama-Nemotron: Efficient Reasoning Models
2025-05-02 · arxiv-v5 · 24 页
Anthropic System CardClaude Sonnet 4 and Opus 4 System Card
2025-05 · 123 页
Microsoft / Phi / MAI Technical ReportPhi-4-reasoning Technical Report
2025-04-30 · arxiv-v1 · 33 页
DeepSeek Technical ReportDeepSeek-Prover-V2 Technical Report
2025-04-30 · 39 页
Moonshot AI / Kimi Technical ReportKimi-Audio Technical Report
2025-04-25 · arxiv-v1 · 26 页
Meta Technical ReportPerception Encoder: The best visual embeddings are not at the output of the network
2025-04-17 · arxiv-v2 · 44 页
OpenAI System CardOpenAI o3 and o4-mini System Card
2025-04-16 · 33 页
ByteDance / Seed Technical ReportSeedream 3.0 Technical Report
2025-04-15 · arxiv-v3 · 22 页
Moonshot AI / Kimi Technical ReportKimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning
2025-04-15 · arxiv-v1 · 24 页
ByteDance / Seed Technical ReportSeed1.5-Thinking: Advancing Superb Reasoning Models with Reinforcement Learning
2025-04-10 · arxiv-v3 · 19 页
Moonshot AI / Kimi Technical ReportKimi-VL Technical Report
2025-04-10 · arxiv-v3 · 24 页
Google / DeepMind Technical ReportTxGemma: Efficient and Agentic LLMs for Therapeutics
2025-04-08 · arxiv-v1 · 58 页
Google / DeepMind Technical ReportEncoder-Decoder Gemma: Improving the Quality-Efficiency Trade-Off via Adaptation
2025-04-08 · arxiv-v1 · 12 页
Amazon / Nova Technical ReportAmazon Nova Sonic: Technical Report and Model Card
2025-04-08 · 11 页
NVIDIA Technical ReportNemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models
2025-04-04 · arxiv-v4 · 35 页
Cohere Technical ReportCommand A: An Enterprise-Ready Large Language Model
2025-04-01 · arxiv-v2 · 55 页
Alibaba / Qwen / Wan Technical ReportWan: Open and Advanced Large-Scale Video Generative Models
2025-03-26 · arxiv-v2 · 60 页
Alibaba / Qwen / Wan Technical ReportQwen2.5-Omni Technical Report
2025-03-26 · arxiv-v1 · 20 页
Google / DeepMind Technical ReportGemma 3 Technical Report
2025-03-25 · arxiv-v1 · 25 页
Google / DeepMind Technical ReportGemini Robotics: Bringing AI into the Physical World
2025-03-25 · arxiv-v1 · 64 页
OpenAI System CardAddendum to GPT-4o System Card: Native image generation
2025-03-25 · 13 页
NVIDIA Technical ReportGR00T N1: An Open Foundation Model for Generalist Humanoid Robots
2025-03-18 · arxiv-v2 · 36 页
NVIDIA Technical ReportCosmos-Reason1: From Physical Common Sense To Embodied Reasoning
2025-03-18 · arxiv-v3 · 36 页
StepFun Technical ReportStep-Video-TI2V Technical Report: A State-of-the-Art Text-Driven Image-to-Video Generation Model
2025-03-14 · arxiv-v1 · 7 页
Alibaba / Qwen / Wan Technical ReportVACE: All-in-One Video Creation and Editing
2025-03-10 · arxiv-v2 · 17 页
ByteDance / Seed Technical ReportSeedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model
2025-03-10 · arxiv-v1 · 33 页
Microsoft / Phi / MAI Technical ReportPhi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
2025-03-03 · arxiv-v2 · 39 页
OpenAI System CardOpenAI GPT-4.5 System Card
2025-02-27 · 31 页
OpenAI System CardDeep Research System Card
2025-02-25 · 36 页
Moonshot AI / Kimi Technical ReportMuon is Scalable for LLM Training
2025-02-24 · arxiv-v1 · 19 页
Alibaba / Qwen / Wan Technical ReportQwen2.5-VL Technical Report
2025-02-19 · arxiv-v1 · 23 页
StepFun Technical ReportStep-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
2025-02-17 · arxiv-v2 · 25 页
Anthropic System CardClaude Sonnet 3.7 System Card
2025-02 · 43 页
OpenAI System CardOpenAI o3-mini System Card
2025-01-31 · 37 页
DeepSeek Technical ReportJanus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
2025-01-29 · arxiv-v1 · 13 页
OpenAI System CardOperator System Card
2025-01-23 · 17 页
Moonshot AI / Kimi Technical ReportKimi k1.5: Scaling Reinforcement Learning with LLMs
2025-01-22 · arxiv-v4 · 25 页
DeepSeek Technical ReportDeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
2025-01-22 · arxiv-v2 · 86 页
Tencent / Hunyuan Technical ReportHunyuan3D 2.0: Scaling Diffusion Models for High Resolution Textured 3D Assets Generation
2025-01-21 · arxiv-v5 · 28 页
MiniMax Technical ReportMiniMax-01: Scaling Foundation Models with Lightning Attention
2025-01-14 · arxiv-v1 · 68 页
NVIDIA Technical ReportCosmos World Foundation Model Platform for Physical AI
2025-01-07 · arxiv-v3 · 75 页
DeepSeek Technical ReportDeepSeek-V3 Technical Report
2024-12-27 · arxiv-v2 · 53 页
Alibaba / Qwen / Wan Technical ReportQwen2.5 Technical Report
2024-12-19 · arxiv-v2 · 26 页
DeepSeek Technical ReportDeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
2024-12-13 · arxiv-v1 · 28 页
Microsoft / Phi / MAI Technical ReportPhi-4 Technical Report
2024-12-12 · arxiv-v1 · 36 页
Meta Technical ReportLarge Concept Models: Language Modeling in a Sentence Representation Space
2024-12-11 · arxiv-v2 · 49 页
OpenAI System CardOpenAI o1 System Card
2024-12-05 · 52 页
Cohere Technical ReportAya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier
2024-12-05 · arxiv-v1 · 19 页
Google / DeepMind Technical ReportPaliGemma 2: A Family of Versatile VLMs for Transfer
2024-12-04 · arxiv-v1 · 31 页
Amazon / Nova Technical ReportThe Amazon Nova Family of Models: Technical Report and Model Card
2024-12-03 · 48 页
Tencent / Hunyuan Technical ReportHunyuanVideo: A Systematic Framework For Large Video Generative Models
2024-12-03 · arxiv-v6 · 35 页
智谱 / Z.ai Technical ReportGLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot
2024-12-03 · arxiv-v1 · 14 页
Meta Technical ReportLlama Guard 3-1B-INT4: Compact and Efficient Safeguard for Human-AI Conversations
2024-11-18 · arxiv-v1 · 9 页
DeepSeek Technical ReportJanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation
2024-11-12 · arxiv-v2 · 25 页
Tencent / Hunyuan Technical ReportHunyuan-Large: An Open-Source MoE Model with 52 Billion Activated Parameters by Tencent
2024-11-04 · arxiv-v3 · 18 页
Meta Technical ReportMovie Gen: A Cast of Media Foundation Models
2024-10-17 · arxiv-v2 · 96 页
DeepSeek Technical ReportJanus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation
2024-10-17 · arxiv-v1 · 24 页
Mistral AI Technical ReportPixtral 12B
2024-10-09 · arxiv-v2 · 24 页
Anthropic Model CardModel Card Addendum: Claude 3.5 Haiku and Upgraded Claude 3.5 Sonnet
2024-10 · 14 页
Alibaba / Qwen / Wan Technical ReportQwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
2024-09-18 · arxiv-v1 · 39 页
Alibaba / Qwen / Wan Technical ReportQwen2.5-Coder Technical Report
2024-09-18 · arxiv-v3 · 32 页
Alibaba / Qwen / Wan Technical ReportQwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
2024-09-18 · arxiv-v2 · 52 页
OpenAI System CardOpenAI o1-preview and o1-mini System Card
2024-09-12 · 43 页
智谱 / Z.ai Technical ReportCogVLM2: Visual Language Models for Image and Video Understanding
2024-08-29 · arxiv-v1 · 27 页
DeepSeek Technical ReportDeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
2024-08-15 · arxiv-v1 · 28 页
智谱 / Z.ai Technical ReportCogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
2024-08-12 · arxiv-v3 · 30 页
OpenAI System CardGPT-4o System Card
2024-08-08 · 33 页
Meta Technical ReportSAM 2: Segment Anything in Images and Videos
2024-08-01 · arxiv-v2 · 42 页
Meta Technical ReportThe Llama 3 Herd of Models
2024-07-31 · arxiv-v3 · 92 页
Alibaba / Qwen / Wan Technical ReportQwen2-Audio Technical Report
2024-07-15 · arxiv-v1 · 16 页
Alibaba / Qwen / Wan Technical ReportQwen2 Technical Report
2024-07-15 · arxiv-v4 · 26 页
Google / DeepMind Technical ReportPaliGemma: A versatile 3B VLM for transfer
2024-07-10 · arxiv-v2 · 59 页
Google / DeepMind Technical ReportGemma 2 Technical Report
2024-06-27 · 21 页
智谱 / Z.ai Technical ReportChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
2024-06-18 · arxiv-v2 · 19 页
NVIDIA Technical ReportNemotron-4 340B Technical Report
2024-06-17 · arxiv-v2 · 34 页
DeepSeek Technical ReportDeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
2024-06-17 · arxiv-v1 · 19 页
Google / DeepMind Technical ReportCodeGemma: Open Code Models Based on Gemma
2024-06-17 · arxiv-v2 · 11 页
Anthropic Model CardClaude Sonnet 3.5 Model Card
2024-06 · 8 页
DeepSeek Technical ReportDeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
2024-05-23 · arxiv-v1 · 17 页
Cohere Technical ReportAya 23: Open Weight Releases to Further Multilingual Progress
2024-05-23 · arxiv-v2 · 27 页
Meta Technical ReportChameleon: Mixed-Modal Early-Fusion Foundation Models
2024-05-16 · arxiv-v2 · 27 页
Tencent / Hunyuan Technical ReportHunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
2024-05-14 · arxiv-v1 · 25 页
智谱 / Z.ai Technical ReportInf-DiT: Upsampling Any-Resolution Image with Memory-Efficient Diffusion Transformer
2024-05-07 · arxiv-v2 · 24 页
DeepSeek Technical ReportDeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
2024-05-07 · arxiv-v5 · 52 页
Microsoft / Phi / MAI Technical ReportPhi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
2024-04-22 · arxiv-v4 · 24 页
Google / DeepMind Technical ReportGemma: Open Models Based on Gemini Research and Technology
2024-03-13 · arxiv-v4 · 17 页
Google / DeepMind Technical ReportGemini 1.5 Technical Report
2024-03-08 · 154 页
DeepSeek Technical ReportDeepSeek-VL: Towards Real-World Vision-Language Understanding
2024-03-08 · arxiv-v2 · 33 页
智谱 / Z.ai Technical ReportCogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion
2024-03-08 · arxiv-v1 · 21 页
Anthropic Model CardClaude 3 Model Card
2024-03 · 64 页
NVIDIA Technical ReportNemotron-4 15B Technical Report
2024-02-26 · arxiv-v2 · 16 页
Cohere Technical ReportAya Model: An Instruction Finetuned Open-Access Multilingual Language Model
2024-02-12 · arxiv-v1 · 118 页
智谱 / Z.ai Technical ReportCogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning
2024-02-06 · arxiv-v3 · 21 页
DeepSeek Technical ReportDeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
2024-02-05 · arxiv-v3 · 30 页
DeepSeek Technical ReportDeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
2024-01-25 · arxiv-v2 · 23 页
DeepSeek Technical ReportDeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
2024-01-11 · arxiv-v1 · 33 页
Mistral AI Technical ReportMixtral of Experts
2024-01-08 · arxiv-v1 · 13 页
DeepSeek Technical ReportDeepSeek LLM: Scaling Open-Source Language Models with Longtermism
2024-01-05 · arxiv-v1 · 48 页
Google / DeepMind Technical ReportGemini 1.0 Technical Report
2023-12-19 · 90 页
智谱 / Z.ai Technical ReportCogAgent: A Visual Language Model for GUI Agents
2023-12-14 · arxiv-v3 · 27 页
Meta Technical ReportSeamless: Multilingual Expressive and Streaming Speech Translation
2023-12-08 · arxiv-v1 · 145 页
Meta Technical ReportLlama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
2023-12-07 · arxiv-v1 · 15 页
Google / DeepMind Technical ReportAlphaCode 2 Technical Report
2023-12-06 · 7 页
Alibaba / Qwen / Wan Technical ReportQwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
2023-11-14 · arxiv-v2 · 18 页
智谱 / Z.ai Technical ReportCogVLM: Visual Expert for Pretrained Language Models
2023-11-06 · arxiv-v2 · 17 页
Google / DeepMind Technical ReportPaLI-3 Vision Language Models: Smaller, Faster, Stronger
2023-10-13 · arxiv-v2 · 16 页
Mistral AI Technical ReportMistral 7B
2023-10-10 · arxiv-v1 · 9 页
Alibaba / Qwen / Wan Technical ReportQwen Technical Report
2023-09-28 · arxiv-v1 · 59 页
OpenAI System CardGPT-4V(ision) System Card
2023-09-25 · 18 页
Microsoft / Phi / MAI Technical ReportTextbooks Are All You Need II: phi-1.5 technical report
2023-09-11 · arxiv-v1 · 16 页
Alibaba / Qwen / Wan Technical ReportQwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
2023-08-24 · arxiv-v3 · 24 页
Meta Technical ReportCode Llama: Open Foundation Models for Code
2023-08-24 · arxiv-v3 · 48 页
Meta Technical ReportSeamlessM4T: Massively Multilingual & Multimodal Machine Translation
2023-08-22 · arxiv-v3 · 111 页
Google / DeepMind Technical ReportRT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
2023-07-28 · arxiv-v1 · 26 页
Meta Technical ReportLlama 2: Open Foundation and Fine-Tuned Chat Models
2023-07-18 · arxiv-v2 · 77 页
Anthropic Model CardClaude 2 Model Card
2023-07 · 16 页
Anthropic Model CardClaude 2 Model Card
2023-07 · 14 页
Microsoft / Phi / MAI Technical ReportTextbooks Are All You Need
2023-06-20 · arxiv-v2 · 26 页
Meta Technical ReportSimple and Controllable Music Generation
2023-06-08 · arxiv-v3 · 17 页
Meta Technical ReportScaling Speech Technology to 1,000+ Languages
2023-05-22 · arxiv-v1 · 41 页
Google / DeepMind Technical ReportPaLM 2 Technical Report
2023-05-17 · arxiv-v3 · 93 页
Meta Technical ReportDINOv2: Learning Robust Visual Features without Supervision
2023-04-14 · arxiv-v2 · 32 页
Meta Technical ReportSegment Anything
2023-04-05 · arxiv-v1 · 30 页
智谱 / Z.ai Technical ReportCodeGeeX: A Pre-Trained Model for Code Generation with Multilingual Benchmarking on HumanEval-X
2023-03-30 · arxiv-v2 · 30 页
OpenAI Technical ReportGPT-4 Technical Report
2023-03-15 · arxiv-v6 · 100 页
OpenAI System CardGPT-4 System Card
2023-03-14 · 60 页
Google / DeepMind Technical ReportPaLM-E: An Embodied Multimodal Language Model
2023-03-06 · arxiv-v1 · 18 页
Meta Technical ReportLLaMA: Open and Efficient Foundation Language Models
2023-02-27 · arxiv-v1 · 27 页
Google / DeepMind Technical ReportMusicLM: Generating Music From Text
2023-01-26 · arxiv-v1 · 15 页
OpenAI Technical ReportRobust Speech Recognition via Large-Scale Weak Supervision
2022-12-06 · arxiv-v1 · 28 页
Meta Technical ReportGalactica: A Large Language Model for Science
2022-11-16 · arxiv-v1 · 58 页
Google / DeepMind Technical ReportImagen Video: High Definition Video Generation with Diffusion Models
2022-10-05 · arxiv-v1 · 18 页
智谱 / Z.ai Technical ReportGLM-130B: An Open Bilingual Pre-trained Model
2022-10-05 · arxiv-v2 · 56 页
Google / DeepMind Technical ReportPaLI: A Jointly-Scaled Multilingual Language-Image Model
2022-09-14 · arxiv-v4 · 33 页
Meta Technical ReportNo Language Left Behind: Scaling Human-Centered Machine Translation
2022-07-11 · arxiv-v3 · 192 页
Google / DeepMind Technical ReportScaling Autoregressive Models for Content-Rich Text-to-Image Generation
2022-06-22 · arxiv-v1 · 49 页
智谱 / Z.ai Technical ReportCogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
2022-05-29 · arxiv-v1 · 15 页
Google / DeepMind Technical ReportPhotorealistic Text-to-Image Diffusion Models with Deep Language Understanding
2022-05-23 · arxiv-v1 · 46 页
Google / DeepMind Technical ReportUL2: Unifying Language Learning Paradigms
2022-05-10 · arxiv-v3 · 39 页
Meta Technical ReportOPT: Open Pre-trained Transformer Language Models
2022-05-02 · arxiv-v4 · 30 页
Google / DeepMind Technical ReportFlamingo: a Visual Language Model for Few-Shot Learning
2022-04-29 · arxiv-v2 · 54 页
智谱 / Z.ai Technical ReportCogView2: Faster and Better Text-to-Image Generation via Hierarchical Transformers
2022-04-28 · arxiv-v2 · 15 页
OpenAI Technical ReportHierarchical Text-Conditional Image Generation with CLIP Latents
2022-04-13 · arxiv-v1 · 27 页
Google / DeepMind Technical ReportPaLM: Scaling Language Modeling with Pathways
2022-04-05 · arxiv-v5 · 87 页
Google / DeepMind Technical ReportTraining Compute-Optimal Large Language Models
2022-03-29 · arxiv-v1 · 36 页
OpenAI Technical ReportTraining language models to follow instructions with human feedback
2022-03-04 · arxiv-v1 · 68 页
Google / DeepMind Technical ReportCompetition-Level Code Generation with AlphaCode
2022-02-08 · arxiv-v1 · 74 页
Google / DeepMind Technical ReportLaMDA: Language Models for Dialog Applications
2022-01-20 · arxiv-v3 · 47 页
Google / DeepMind Model CardVeo 3.1 Lite Model Card
日期待确认 · 4 页
Google / DeepMind Model CardVeo 3 Model Card
日期待确认 · 6 页
Google / DeepMind Technical ReportShieldGemma 1 Technical Report
日期待确认 · 11 页
ByteDance / Seed Model CardSeed2.1 Model Card: Agentic Intelligence for Productivity
日期待确认 · 73 页
ByteDance / Seed Model CardSeed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity
日期待确认 · 78 页
Google / DeepMind Technical ReportRecurrentGemma: Moving Past Transformers for Efficient Open Language Models
日期待确认 · 6 页
NVIDIA Technical ReportNVIDIA Nemotron Nano V2 VL
日期待确认 · 29 页
NVIDIA Technical ReportNVIDIA Nemotron 3 Ultra Technical Report
日期待确认 · 65 页
NVIDIA Technical ReportNVIDIA Nemotron 3 Super Technical Report
日期待确认 · 51 页
Meta System CardMuse Spark 1.1 Evaluation Report
日期待确认 · 112 页
Microsoft / Phi / MAI Model CardMAI-Voice-2.1-Flash Model Card
日期待确认 · 4 页
Microsoft / Phi / MAI Model CardMAI-Voice-2.1 Model Card
日期待确认 · 4 页
Microsoft / Phi / MAI Model CardMAI-Voice-2 Model Card
日期待确认 · 3 页
Microsoft / Phi / MAI Model CardMAI-Transcribe-2-Streaming Model Card
日期待确认 · 4 页
Microsoft / Phi / MAI Model CardMAI-Transcribe-2 Model Card
日期待确认 · 8 页
Microsoft / Phi / MAI Model CardMAI-Image-2.6 / MAI-Image-2.6-Flash Model Card
日期待确认 · 4 页
Microsoft / Phi / MAI Model CardMAI-Cyber-1-Flash Model Card
日期待确认 · 7 页
Microsoft / Phi / MAI Model CardMAI-Code-1.1-Flash Model Card
日期待确认 · 6 页
Google / DeepMind Model CardLyria 3.5 Model Card
日期待确认 · 4 页
Google / DeepMind Model CardLyria 3 Model Card
日期待确认 · 5 页
Moonshot AI / Kimi Technical ReportKimi K3 Technical Report
日期待确认 · 47 页
OpenAI Technical ReportImproving Image Generation with Better Captions
日期待确认 · 19 页
Google / DeepMind Model CardImagen 4 Model Card
日期待确认 · 7 页
xAI Model CardGrok Code Fast 1 Model Card
日期待确认 · 6 页
xAI Model CardGrok 4 Model Card
日期待确认 · 8 页
xAI Model CardGrok 4 Fast Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini Robotics-ER 2 Model Card
日期待确认 · 5 页
Google / DeepMind Model CardGemini Robotics-ER 1.6 Model Card
日期待确认 · 4 页
Google / DeepMind Model CardGemini Robotics On-Device Model Card
日期待确认 · 5 页
Google / DeepMind Model CardGemini Robotics On-Device 2 Model Card
日期待确认 · 5 页
Google / DeepMind Technical ReportGemini Robotics 1.5 Technical Report
日期待确认 · 62 页
Google / DeepMind Model CardGemini Omni Flash Model Card
日期待确认 · 4 页
Google / DeepMind Model CardGemini 3.8 Flash Model Card
日期待确认 · 8 页
Google / DeepMind Model CardGemini 3.8 Audio (Live, Live Extended Thinking, Flash TTS, Flash-Lite TTS) Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.7 Flash Model Card
日期待确认 · 9 页
Google / DeepMind Model CardGemini 3.6 Flash Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.5 Flash-Lite Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.5 Flash Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.5 Audio (Live Translate, Transcribe, Transcribe Live) Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.1 Pro Model Card
日期待确认 · 9 页
Google / DeepMind Model CardGemini 3.1 Flash-Lite Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.1 Flash-Lite Image Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.1 Flash Image Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3.1 Flash Audio (Flash Live, TTS) Model Card
日期待确认 · 5 页
Google / DeepMind Model CardGemini 3 Pro Model Card
日期待确认 · 10 页
Google / DeepMind Model CardGemini 3 Pro Image Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 3 Flash Model Card
日期待确认 · 6 页
Google / DeepMind Model CardGemini 2.5 Pro Model Card
日期待确认 · 21 页
Google / DeepMind Model CardGemini 2.5 Flash-Lite Model Card
日期待确认 · 9 页
Google / DeepMind Model CardGemini 2.5 Flash and Gemini 2.5 Flash Image Model Card
日期待确认 · 11 页
Google / DeepMind Model CardGemini 2.5 Deep Think Model Card
日期待确认 · 20 页
Google / DeepMind Model CardGemini 2.5 Computer Use Model Card
日期待确认 · 5 页
Google / DeepMind Model CardGemini 2.0 Flash-Lite Model Card
日期待确认 · 7 页
Google / DeepMind Model CardGemini 2.0 Flash Model Card
日期待确认 · 7 页
DeepSeek Technical ReportDeepSeek-V3.2-Exp: Boosting Long-Context Efficiency with DeepSeek Sparse Attention
日期待确认 · 6 页
Amazon / Nova System CardAWS AI Service Cards: Amazon Nova Act
日期待确认 · 16 页
OpenAIGPT-6.1 Sol
通用与推理 / 编程
OpenAIGPT-6 Astra
通用与推理 / 编程
OpenAIChatGPT Images 2.5 / Sunburst / Flare
图像生成
OpenAIGPT-Rosalind-5.5
科学研究
AnthropicClaude Opus 5.5
通用与推理 / 编程
AnthropicClaude Fable 5.1 / Mythos 5.1
通用与推理 / 编程 / 安全
Google / DeepMindGemini 3.8 Flash
通用与推理 / 编程 / 视觉理解
Google / DeepMindGemini 3.1 Pro
通用与推理 / 编程 / 视觉理解
Google / DeepMindLyria 3.5
音乐生成
Google / DeepMindGemma 4
通用与推理 / 视觉理解
Google / DeepMindMedGemma 1.5
医疗
Google / DeepMindTranslateGemma
翻译
MetaMuse Spark 1.3
通用与推理 / 编程 / 视觉理解
MetaMuse Glimmer 30B
通用与推理
MetaMuse Image
图像生成
MetaMuse Video
视频生成
MetaMuse Voice Transcribe
语音识别
MetaLlama 4 Scout / Maverick
通用与推理 / 视觉理解
xAIGrok 4.7
通用与推理 / 编程
xAIGrok Imagine Image 2.0
图像生成
xAIGrok Imagine Video 1.5
视频生成
Mistral AIMistral Medium 3.5
通用与推理 / 编程
Mistral AIMistral Large 3
通用与推理 / 视觉理解
Mistral AIVoxtral TTS
语音生成
Mistral AIVoxtral Realtime
实时语音 / 语音识别
Mistral AIShieldstral
安全
Mistral AIRobostral Navigate
机器人
Mistral AIMistral OCR 4.1
文档理解
NVIDIANemotron 3 Super
通用与推理 / 编程
NVIDIANemotron 3 Ultra
通用与推理 / 编程
NVIDIAGR00T N1.7
机器人
DeepSeekDeepSeek-V4.1-Flash
通用与推理 / 视觉理解
Alibaba / Qwen / WanQwen3.8-Flash-Next
通用与推理
Moonshot AI / KimiKimi K3
通用与推理 / 视觉理解
智谱 / Z.aiGLM-5.3
通用与推理 / 编程
MiniMaxMiniMax-M3
通用与推理 / 编程
MiniMaxMiniMax-H3
视频生成
MiniMaxMiniMax Music 3.0
音乐生成
MiniMaxMiniMax Speech 2.8
语音生成
StepFunStep 3.7 Flash
通用与推理
Tencent / HunyuanHy4-preview
通用与推理
Tencent / HunyuanHy-MT2
翻译
ByteDance / SeedSeed2.1
通用与推理 / 视觉理解
ByteDance / SeedSeedance 2.5
视频生成
ByteDance / SeedSeedream 5.0 Pro
图像生成
CohereCommand A+
通用与推理
CohereNorth Mini Code 1.0
编程
CohereNorth Micro Vision Instruct
视觉理解
CohereCohere Transcribe Arabic
语音识别
CohereNorth Small Translate 1.0
翻译
CohereEmbed 5 (pro / fast)
向量嵌入
CohereRerank 4 (pro / fast)
重排序
CohereParse 5
文档理解
Microsoft / Phi / MAIMAI Thinking 1
通用与推理
Microsoft / Phi / MAIMAI Code 1.1 Flash
编程
Microsoft / Phi / MAIMAI Image 2.6 / 2.6 Flash
图像生成
Microsoft / Phi / MAIMAI Cyber 1 Flash
安全
Microsoft / Phi / MAIPhi 4 Reasoning Vision
视觉理解
Amazon / NovaNova 2 Sonic
实时语音
OpenAIGPT-Live-1
实时语音
OpenAIGPT-Transcribe / GPT-Live-Transcribe
语音识别
OpenAIGPT-4o Mini TTS
语音生成
OpenAIGPT-Realtime-Translate
翻译
OpenAIGPT-5.6 Cyber
安全
AnthropicClaude Sonnet 5.5
通用与推理 / 编程
xAIGrok Voice Transcribe 2.0
语音识别
xAIGrok Voice Think Fast 2.0
实时语音
Mistral AICodestral Embed
向量嵌入
Microsoft / Phi / MAIMAI-Transcribe-2-Streaming
语音识别
Microsoft / Phi / MAIMAI-Voice-2.1 / Flash
语音生成
Amazon / NovaNova 2 Lite / Pro Preview
通用与推理
Amazon / NovaNova 2 Omni Preview
图像生成
Amazon / NovaNova Multimodal Embeddings
向量嵌入
DeepSeekDeepSeek-OCR-2
文档理解
Alibaba / Qwen / WanQwen3.8-Omni-Flash
通用与推理 / 视觉理解
Alibaba / Qwen / WanQwen3-Coder-Next
编程
Alibaba / Qwen / WanQwen-Image-2.1-Pro
图像生成
Alibaba / Qwen / WanWan3.0-Video-Prime
视频生成
Alibaba / Qwen / WanQwen-Audio-3.0-ASR
语音识别
Alibaba / Qwen / WanQwen-Audio-3.0-TTS-Plus / Flash
语音生成
Alibaba / Qwen / WanQwen3.8-Omni-Flash-Realtime
实时语音
Alibaba / Qwen / WanQwen-MT-Uni
翻译
Moonshot AI / KimiKimi K2.8 Preview
编程
Moonshot AI / KimiKimi-Audio
语音理解
智谱 / Z.aiGLM-5.3-FlashX
通用与推理 / 视觉理解
智谱 / Z.aiGLM-Image
图像生成
智谱 / Z.aiGLM-ASR-Nano
语音识别
智谱 / Z.aiGLM-TTS
语音生成
智谱 / Z.aiGLM-OCR
文档理解
StepFunStep 5 Preview
通用与推理 / 视觉理解
StepFunNextStep-1.1
图像生成
StepFunStep-Video-TI2V
视频生成
StepFunStep1X-3D
3D 生成
StepFunStepAudio 3 ASR Max
语音识别
StepFunStepAudio 3 TTS
语音生成
StepFunStepAudio 3 Realtime
实时语音
StepFunStepAudio 3 Music
音乐生成
Tencent / HunyuanHunyuanImage 3.0
图像生成
Tencent / HunyuanHunyuanVideo 1.5
视频生成
Tencent / HunyuanWeVisDoc-4B / 2B
文档理解
Tencent / HunyuanHunyuan3D 3.1
3D 生成
ByteDance / SeedSeedRealtime
实时语音
ByteDance / SeedSeed Audio 1.0
语音理解
Google / DeepMindGemini 4 Argon (limited release)
通用与推理 / 编程 / 视觉理解 / 安全
Google / DeepMindNano Banana 2 Lite / Gemini 3.1 Flash-Lite Image
图像生成
Google / DeepMindGemini Omni 1.1 Flash
视频生成
Google / DeepMindGemini 3.8 Live / Live Extended Thinking
实时语音
Google / DeepMindGemini 3.8 Flash TTS / Flash-Lite TTS
语音生成
Google / DeepMindGemini 3.5 Transcribe / Transcribe Live
语音识别
Google / DeepMindGemini Embedding 2
向量嵌入
Google / DeepMindGemini Robotics 2 / ER 2 / On-Device 2
机器人
Google / DeepMindGenie 3
世界模型
Google / DeepMindWeatherNext 3
科学研究
MetaSAM 3D Objects / Body
3D 生成
NVIDIANemotron 3.5 Lightning
通用与推理
NVIDIANemotron Labs 3 Competitive Coding 550B-A55B
编程
NVIDIANemotron 3 Nano Omni
视觉理解
NVIDIANemotron Parse 2.0
文档理解
NVIDIACosmos 3 Edge
世界模型
NVIDIANemotron 3.5 ASR Streaming 0.6B
语音识别
NVIDIANemotron Labs VoiceChat 11B
实时语音
NVIDIAMagpieTTS Multilingual 357M v2607
语音生成
NVIDIANemotron 3 Embed 1B / 8B
向量嵌入
Alibaba / Qwen / WanQwen3.8-Max
通用与推理 / 编程 / 视觉理解
DeepSeekDeepSeek-V4-Pro
通用与推理 / 编程