🚀 中国大模型动态 · 全部

共 24 条 · 更新于 2026-08-05 20:00 UTC · ← 返回首页

Mastodon·LLM📦 DeepSeek V4 Flash 0731 lands on the leaderboard with open weights: 1M context📦 DeepSeek V4 Flash 0731 lands on the leaderboard with open weights: 1M context at $0.09 in / $0.18 out. A new option for long-context tasks without t…2小时前 Reddit r/LocalLLaMADeepSeek V4 Flash 0731 at 10–17 t/s (nothink) on MacBook M5 Pro **64GB***, partly via SSD streamingInspired by a post from u/giveen I motivated claude (no patinence on my side to work through everything myself) to help me get DS running on my MacBoo…2小时前 Mastodon·AILe dernier analog nowhere. :) http:// analognowhere.com/_/cxcgel/ # aiLe dernier analog nowhere. :) http:// analognowhere.com/_/cxcgel/ # ai2小时前 Reddit r/LocalLLaMAGiven the MiniMax H3 LoRAs Debacle - Some Important Context for Censorship enforcement and laws in China*I felt the need to write this post because it seems like very few people on this sub are aware of Chinese laws and how they're enforced, so here's an…5小时前 IT之家消息称阶跃星辰内部确立大模型与智能体终端两条战略线,手机业务独立运营、将进入海外市场IT之家 8 月 5 日消息,据虎嗅消息, 阶跃星辰内部正在按照两条战略线推进:一条仍以大模型为核心,另一条则以智能体终端为核心, 目前手机只是初步探索。 其中,以大模型为核心的业务仍以上海阶跃星辰智能科技股份有限公司这一原有主体运营; 手机业务则被放入一家新公司独立运营 。 据悉,阶跃正在搭建海外…6小时前 Reddit r/LocalLLaMAMoE CPU-offload benchmark on Deepseek V4/Gemma4/Qwen/GPT-OSS — TensorSharp vs llama.cppTensorSharp's MoE CPU-offload feature has been merged into main. Here is the parameters description of this feature: Mixture-of-Experts CPU offload: -…6小时前 Reddit r/LocalLLaMAMiniMax issueshttps://www.reddit.com/r/StableDiffusion/s/HrU7odaJe6 I think this is more important that all the political stuff you share here submitted by /u/jacek…7小时前 Reddit r/LocalLLaMAQwen Developers' responses from their recent Twitter/X AMAQuestions & Responses(in BOLD ) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folk…8小时前 雷锋网MiniMax H3 视频模型登顶开源社区第一,定义视频模型领域“斩杀线”本周,MiniMax 正式开源新一代通用多模态生成模型 MiniMax H3,在 Artificial Analysis 视频编辑评测榜、Arena 图生视频排行榜中均位列全球第一,并获得 Stable Diffusion 创始人 Emad Mostaque、硅谷投资机构 a16z 合伙人 Just…8小时前 Reddit r/LocalLLaMAI updated my localy run benchmark with DeepSeek V4 Flash 0731It's the purple cluster on the top left (the good corner...) I'm running the MXFP4 version from Bartoswski with Dspark at 1K t/s prefill and 90 t/s ge…9小时前 Reddit r/LocalLLaMAQwen3-TTS voice cloning is now in mainline llama.cpp — the old demo finally became real supportPeople may remember the Qwen3-TTS llama.cpp demo from a few months ago. That PR said it probably wouldn’t be merged because llama.cpp was missing some…12小时前 爱范儿8GB 内存也能跑 Kimi K3?2026 本地部署大模型配置全指南本地 AI 设备会越来越贵吗 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。13小时前 Reddit r/LocalLLaMAKimi K3 full model running on 16x GB10 cluster at 20+tpsKimi K3 full model running on 16x GB10 cluster at 20+tps average (llama-benchy coherent corpus) 38tps peak, 750tps prefill. This is the first run of f…昨天 Simon WillisonPipeNetwork/minimax-h3-mlxPipeNetwork/minimax-h3-mlx MiniMax released MiniMax-H3 two days ago - they describe it as a "a general-purpose, omni-modal generative system", which i…昨天 Reddit r/LocalLLaMAHas anyone tried Mach-1 Additive? 95% of performance of Qwen 3.6 35B while being 10x smallerWhy nobody is talking about this? Seems pretty significant to the community submitted by /u/MuzafferMahi [link] [comments]昨天 雷锋网发布当日,海外主流AI平台纷纷接入阿里Qwen3.88月3日,阿里巴巴正式发布新一代基座大模型Qwen3.8,编程与Agent能力大幅提升,整体性能位居全球大模型第一梯队。发布当日,OpenRouter、OpenCode、Hermes Agent、Command Code、Vercel、Novita、Charm、DeepInfra等多家海外主流API…昨天 极客公园飞书并入豆包:字节跳动的一次目标重塑头图来源:视觉中国 7 月 30 日,字节跳动下发内部邮件,完成了近年来力度最大的一次 ToB 业务架构调整:飞书产品团队与豆包产品团队整合,成立新的豆包产品团队,由豆包负责人赵祺统筹,飞书负责人谢欣向其汇报;飞书商业化团队则与火山引擎整合,成立新的 ToB GTM 组织「创造力服务平台」,由火山引…昨天 少数派派早报:MiniMax H3 开源、Qwen3.8-Max 发布等少数派的近期动态投稿征文赢好礼!「角落新声」征文活动距离结束还有一周。火速投稿新一季少数派会员启航,更新权益,更多惊喜,还有实体纪念卡。点击了解你可能错过的文章角落新声|两平米、两个角落,安放两个自己 ... 查看全文昨天 极客公园传新 iPhone 最高涨价 1350 元;阿里发 Qwen3.8,2.4 万亿参数;DuckDuckGo 推「反科技」太阳镜白宫拟召集 OpenAI 和 Anthropic 等巨头,讨论 AI 模型安全测试框架 据五位知情人士透露,特朗普政府已邀请 OpenAI、谷歌、Anthropic 等大型科技企业工作人员于周二前往白宫,审议人工智能监管框架最终版本。该框架将设立一套自愿性机制,要求人工智能实验室在向合作方及公众推出…昨天 爱范儿实测千问办公:阿里这个 AI 专家,真的能帮你带货Agent 应用接下来比的,或许不只是模型有多聪明。 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。2天前 爱范儿AI 改变打工人之后,千问办公盯上了企业 Agent 化办公 Agent ,怎么提升企业组织的办公效率 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。2天前 爱范儿DeepSeek 给大模型划出的「斩杀线」,斩的到底是什么斩掉了大量缺乏明确不可替代性的中间模型 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。2天前 Simon Willisondeepseek-ai/DeepSeek-V4-Flash-0731deepseek-ai/DeepSeek-V4-Flash-0731 The latest release in DeepSeek's V4 family, "with substantially enhanced agentic capabilities". It's 304 billion pa…4天前 极客公园宇树科技 8 月 10 日开启 IPO 申购;字节 ToB 变阵:飞书豆包火山组织调整;马斯克称人类 5-7 年登陆火星 | 极客早知道宇树科技:初步询价日为 8 月 5 日,网下申购日为 8 月 10 日 7 月 30 日宇树科技公告称,公司首次公开发行股票并在科创板上市,本次发行采用战略配售、网下发行与网上发行相结合的方式进行。 本次拟公开发行股份 4044.6434 万股,占发行后总股本的 10%,发行后总股本为 40446.…5天前