GPT-6 家族的模范指南
原文: A model guide for the GPT-6 family
OpenAI published guidance for startups on choosing among GPT-6 family models, setting reasoning effort, and preparing prompts, tools and workflows for production.
每日彙整官方發布、Hacker News 與 arXiv 的 AI 新聞
本頁標題與摘要由英文離線機器翻譯,可能有誤。請以連結的原文為準。
無權蒸馏模型產品是AI實驗室日益擔心的安全問題。,這是OpenAI自己調查和打亂一場競選的說法。
OpenAI
An 8B open-weights model built specifically for cited scientific reports that teams can run on their own infrastructure.
Ai2 (Allen Institute for AI)
Cloudflare's post on open-weight decision models and an RL fine-tuning platform was the most-upvoted AI story on Hacker News in this edition's 36-hour window.
Hacker News
編輯附註: 摘要是根據各發布者的原文、在 AI 協助下撰寫的草稿,尚待人工審閱。
過去 7 天內發布於公司與研究機構官方部落格的文章。
原文: A model guide for the GPT-6 family
OpenAI published guidance for startups on choosing among GPT-6 family models, setting reasoning effort, and preparing prompts, tools and workflows for production.
原文: The eternal complement
An OpenAI essay argues that advanced AI may matter most by speeding up the routine execution work behind breakthroughs, which could set the pace of economic progress.
原文: How Albertsons Companies is reimagining retail from the inside out
Customer story: grocery retailer Albertsons Companies describes using ChatGPT Enterprise and the OpenAI API to speed up internal work and make shopping easier for customers.
原文: The Den frees up 10-15 hours a week to grow with ChatGPT Work
客戶故事:社會俱樂部 The Den說, ChatGPT 工作將授權申請等工作從一天缩短到幾小時。,每周省下10到15小時
OpenAI says it disrupted a coordinated attempt to extract protected model reasoning through distillation and is strengthening its defenses against such attacks.
原文: Helping small businesses put AI to work
OpenAI與美國的SBDC合作, 以擴展對小企業的實際人工智能訓練與當地支援。,以及小組如何使用AI的報告。
ServiceNow AI describes AutoSynthData, a method for turning an enterprise agent's observed weaknesses into many new, verifiable training tasks that fit the target environment.
原文: Open TTS Leaderboard: Scalable Evaluation for Multilingual Text-to-Speech and Voice Cloning
Hugging Face介紹了 Open TTS 領導板, 使多語文字對語言和語音克隆模型的評估比今天的破碎更標準化,基于竞技場的比對。
NVIDIA presents Kumo Tabular, a tabular foundation model that predicts new rows through in-context learning instead of training a separate gradient-boosted model for each task.
原文: Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents
Multiverse Computing explains its ProvenanceGuard paper, which checks not only whether an MCP agent's claim is true but whether it is attributed to the right source.
原文: Holo4: powering generalist computer-use agents
H Company released Holo4, agentic computer-use models in 27B dense and 35B-A3B mixture-of-experts sizes on its H Models API, plus an updated Holotron4 Nano.
Mistral AI正在開立一個慕尼黑物理中心 AI和工業AI研究,和德國工業的夥伴們合作
原文: NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI
NVIDIA提供64GB DGX Spark, 供開發者在當地建立和運行日益有能力的開放模型和AI代理。
原文: How NVIDIA GPUs Help Accelerate OpenAI’s GPT-6 Astra Ultrafast
NVIDIA說OpenAI的GPT-6 Astra Ultrafast在NVIDIA Blackwell GPU上运行,此模型可以在 OpenAI API 中找到, 並且可以讓符合資格的 ChatGPT Work 和 Codex 使用者使用 。
NVIDIA makes its case for AI-factory returns, citing a cost of roughly $60 million per megawatt and arguing operators need a clear ROI before committing capital.
原文: NVIDIA Opens Applications for 2027–2028 Graduate Fellowships With Awards Up to $60,000
NVIDIA opened applications for its 2027–2028 Graduate Fellowship program, with awards of up to $60,000.
NVIDIA and CoreWeave describe their long co-engineering partnership and an AI-focused cloud covering agentic AI from training to production.
原文: Open-sourcing AstaBrief, the fast report-generation model in Asta
Ai2 released AstaBrief, an 8B open-weights model that writes cited scientific reports; it powers Asta's Fast mode and can be downloaded to run on your own infrastructure.
Ai2 introduced Olmo-core 3, a redesigned and fully open training stack for scaling mixture-of-experts models toward the trillion-parameter range.
原文: New tool lets users repair AI-generated 3D models, then fabricate them just the way they want
MIT researchers present InstructMesh, a tool that produces easy-to-edit designs of everyday objects so both experts and beginners can repair AI-generated 3D models and fabricate them.
原文: MIT Transit Lab to develop an AI platform for public transit agencies
MIT's Transit Lab will build an open-source Public Transit Intelligence Hub, funded with $2.1 million from Google.org, to unify transit monitoring, operations and rider communication.
原文: This game-playing AI is the new champ at Stratego
MIT News報導了一個AI系統,它擊敗了名列前茅的人類Stratego玩家,而且比其他模型更有效率.,研究者們看到策略决策中的用途。
An MIT study of hiring decisions finds that many firms relying on the same algorithm can, in some situations, benefit job seekers.
原文: Who we become when we talk to machines
MIT professor Sherry Turkle's new book, "Artificial Intimacy," critiques chatbots and the antisocial dynamics she argues they encourage.
過去 36 小時內標題與 AI 相關的文章,依分數排序。
原文: DeepSeek Harness Desktop for macOS and Windows
原文: Meta Uses A.I. Data Centers to Avoid Billions in Federal Taxes
原文: FTC is investigating OpenAI, Anthropic and other AI companies over product risks
原文: Vote on which of Hacker News' challenges for AI have been met
原文: AI Makes Me Sad
原文: Identity Management for Agentic AI [pdf] (2025)
原文: Don't be fooled–LLMs don't reason
自動挑選新投稿(優先選擇同時列於 cs.AI、cs.CL、cs.LG 的論文),並非品質排名。
原文: Comedic Fool's Gold: Reward Exploits and Countermeasures in Conversational Humor
We investigate automated rewards for training language models in conversational humor, focusing on reward exploits and countermeasures.
當一個知情的對手分享一個受限的訊號頻道的觀眾,最能保護真理的訊號 就是最能形容真理的訊號
原文: ContractRL: Shielded Group-Relative Policy Optimization for Auditable Tool-Call Repair
Structured tool calls often fail after only a small number of fields violate a schema or an execution contract.
自動拉帶优化可以通过迭代更新其提示而大大改善 LLM 代理,工具介面,從執行回應中控制邏輯 。
Attributing model behavior to synthetic training data requires knowing what produced each training item before estimating what that item caused.
原文: How Divergence Becomes Decision Flips in Compressed Language Models
Compression reports summarize how far a compressed language model moved from the dense one, usually by a KL divergence; a deployment that relies on the dense model's outputs needs to know how many of its decisions changed.
We propose ReHoPER, an inference-only, zero-shot method that improves large language models' reasoning by generating and answering intermediate questions along multiple paths before the final answer.
知識分解旨在將大型語言模型的實際知識轉移到更小的模型,以高效的部署.
原文: Capturing In-Context Learning Dynamics with Task Operators
內文學習(ICL)讓語言模型可以在沒有重點更新的情况下從演示中完成新的工作。
This work presents Diffusion Layer Integrated Gradients (DLIG), a token attribution method for diffusion language models (DLMs) that extends Integrated Gradients (IG~\cite{sundararajan2017axiomatic}) to arbitrary layers and denoising steps.
重放選取者常以格式回應來排序快取的軌道,B. 信心,新鲜度,或回應長度,雖然缓存水平的正确性和下游學者效用是不同的目標。
培训前的干预措施对于协调研究至关重要,因為在訓練前的形狀中 信仰會形成 一個模型如何從後來訓練中概括出來
原文: CARM: Cancellation-Aware Response Masking for LLM Reinforcement Learning
Recent years have witnessed the rapid adoption of reinforcement learning (RL) in large language model (LLM) post-training, with substantial gains in mathematical reasoning and code generation.
原文: Finetuning with Sampling: SFT Learns Better Than You Think
Introducing new capabilities to frontier models has long been the goal of posttraining, which predominantly employs supervised finetuning (SFT) and reinforcement learning (RL) to this end.
頻率調整注意 [Zeris],2026e]通过在學習頻率上用波段通道过滤的內部產品取代Q/K點產品,在標準點產值注意上取得了很大的收益.
未收錄:
更新: 2026-10-02 16:45 UTC