环球老虎财经 on MSN
做一个能自迭代的后训练平台,Mind Lab要让更多企业拥有自己的模型
在 Infra 层面下功夫,正成为一个行业共识。
Google Research open-sources RRSI, letting agents improve prompts, tools and memory without overfitting, lifting ...
On September 29, 2026, Rep. Ro Khanna (D-Calif.) asked Chinese AI companies and US intelligence officials to assess their ...
"My child has been attending a major programming school since elementary school for three years and can now create games ...
2026年春天,机器人界突然‘抖’了一下——不是因为某家科技巨头又砸了百亿美金,而是一支来自香港大学MMLAB实验室的团队,用一套名字拗口、逻辑却极简的系统PhysicalRSI1.0,在全球最硬核的具身智能竞技场RoboDojo上,直接把‘精密操作’这个老大难问题,从4%的成功率,一把拉到了32.50%!更魔幻的是:它干这事,只花了GPT-6Astra一次全量测试成本的1%——也就是不到2000 ...
导读|RSI(递归自我改进)不应只发生在模型参数层:类比人类大脑的不是 LLM,而是整个 Agent。EverMind 开源的 Raven V0.2.0 为此提供了两个核心设计:为 RSI 而设计的 Harness 框架,以及统一编排专业 Agent ...
Empowering individuals with hands-on model training skills delivers far greater long-term value than merely offering ...
最近几个月,后训练成了 AI 行业最集中的议题。 基础模型公司在扩大 RL 的规模;应用公司也在考虑,如何基于真实业务里的任务轨迹、用户反馈构建自己的后训练流程。 它们指向了同一个变化: 模型在真实环境所产生的经验,正在成为下一代智能的原料。
Mint Recursive aims to help businesses develop their own evaluation frameworks, reward mechanisms, and training data so they ...
红板报 on MSN
马卡龙的下注:把模型后训练,做成企业服务
Mind Lab 发布模型后训练与推理平台 Mint Recursive,把 MinT 等内部训练链路开放成托管服务。文章复盘 Macaron V1.1 以 GLM-5.3 为基座、训练四个 2B LoRA 专家的过程,以及发票核对与生成式 UI ...
Das intensive Engagement auf der Infrastrukturebene etabliert sich zunehmend zu einem branchenweiten Konsens – ein zentraler Trend für zukunftsorientierte Unternehmen, die ihre Wettbewerbsfähigkeit du ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results