24小时 AI快讯 实时更新
聚合全球 AI 厂商、模型、API、开源项目与服务状态动态。
LATEST UPDATES
最新获取
2026-09-02
-
The production platform for open-weight AI inference
Run open models in production with full control over performance, cost, and quality. Deploy in minutes, roll out safely, and scale to your SLOs.
Together AI产品动态查看详情 -
Kimi K3 vs Claude Fable 5 on DeepSWE: Cost and Coding
We ran 452 DeepSWE rollouts on Kimi K3 and Claude Fable 5. Fable leads pass@1 by 1.4 points; Kimi K3 wins pass@4 and delivers 2.8x the solves per dollar.
Together AI产品动态查看详情 -
Kimi K3 vs GPT-5.6 Sol on DeepSWE: Cost, Coding, and Routing
We ran 904 DeepSWE rollouts on Kimi K3 and GPT-5.6 Sol. Sol leads pass@1; Kimi K3 wins pass@4 at 2.8x the solves per dollar, and routing between them reaches ~85.6%.
Together AI产品动态查看详情 -
ThunderAgent: 2x Faster Agentic Inference for Synthetic Data Generation at Scale
ThunderAgent is a program-aware scheduler for agentic inference. By treating each agent workflow as a schedulable program, it eliminates KV cache thrashing to deliver more than 2x single-node throughput and near-linear multi-node scaling.
Together AI产品动态查看详情 -
Configuring Dedicated Model Inference
The three-part resource model behind Together AI Dedicated Model Inference—endpoints, deployments, configs—and how capacity-aware routing ties them together.
Together AI产品动态查看详情 -
Together AI announces strategic partnership with Moonshot AI to natively serve Kimi models
Together AI partners with Moonshot AI to natively serve Kimi models, starting with the 2.8T parameter Kimi K3, with day zero access and post-training.
Together AI产品动态查看详情 -
Autoscaling endpoints for LLM inference
GPU utilization can read healthy while your queue backs up, and a new replica takes minutes to warm. Here's how to pick autoscaling metrics, tune scale-up/down windows, and budget for cold starts on dedicated inference.
Together AI产品动态查看详情 -
Kimi K3: the complete developer guide
Kimi K3 is the first open 3T-class model. See how it benchmarks, what it costs, and how to call it on the Together AI API, with copy-paste code examples.
Together AI产品动态查看详情 -
DeepSeek-V4 Flash 0731 vs GPT-5.6 Luna on DeepSWE: Cost and Coding
We ran 900 DeepSWE rollouts on DeepSeek-V4 Flash and GPT-5.6 Luna. Luna leads pass@1 by 14 points; DeepSeek delivers 4.8x the solves per dollar.
Together AI产品动态查看详情 -
Kevin32223112/miki7
Hugging Face模型更新查看详情 -
tritueviet/sow-ternary-mini-identity-relay
Hugging Face模型更新查看详情 -
mradermacher/gemma-4-12b-it-3MPER0RR-abliterated-i1-GGUF
Hugging Face模型更新查看详情 -
mradermacher/Jenzin-Wuang-Nemotron-30B-A3B-BF16-GGUF
Hugging Face模型更新查看详情 -
voicist/nm-fill-1020
Hugging Face模型更新查看详情 -
Toleng/koplak-flash-1.5b
Hugging Face模型更新查看详情 -
sundaycoil/bookmark-manager
Hugging Face模型更新查看详情 -
aixk/FHN-100M-WaveCross
Hugging Face模型更新查看详情 -
zhc12/dynrank-compressed-models
Hugging Face模型更新查看详情 -
Tostibrown/Qwen3.8-27B-5bit-affine-g64
Hugging Face模型更新查看详情 -
reyansh38771/sn97____ringtone-ro____uid142____hk5F6sC
Hugging Face模型更新查看详情