ไกด์ต้นแบบสําหรับครอบครัว GPT-6
ต้นฉบับ: A model guide for the GPT-6 family
OpenAI published guidance for startups on choosing among GPT-6 family models, setting reasoning effort, and preparing prompts, tools and workflows for production.
ข่าว AI รายวันจากแหล่งทางการ Hacker News และ arXiv
ชื่อเรื่องและสรุปในหน้านี้แปลด้วยเครื่องจากภาษาอังกฤษ (ออฟไลน์) และอาจมีข้อผิดพลาด โปรดยึดต้นฉบับที่ลิงก์ไว้
Unauthorized distillation of model outputs is a growing security concern for AI labs; this is OpenAI's own account of detecting and disrupting one campaign.
OpenAI
โมเดลเปิดรุ่น 8B สร้างมาโดยเฉพาะ สําหรับรายงานทางวิทยาศาสตร์ที่อ้างอิงได้ ว่าทีมสามารถทํางานได้บนโครงสร้างพื้นฐานของตัวเอง
Ai2 (Allen Institute for AI)
Cloudflare's post on open-weight decision models and an RL fine-tuning platform was the most-upvoted AI story on Hacker News in this edition's 36-hour window.
Hacker News
หมายเหตุบรรณาธิการ: สรุปเหล่านี้ร่างขึ้นโดยใช้ AI ช่วยจากข้อความของผู้เผยแพร่แต่ละราย และยังรอการตรวจทานโดยมนุษย์
โพสต์จากบล็อกทางการของบริษัทและห้องวิจัยที่เผยแพร่ใน 7 วันที่ผ่านมา
ต้นฉบับ: A model guide for the GPT-6 family
OpenAI published guidance for startups on choosing among GPT-6 family models, setting reasoning effort, and preparing prompts, tools and workflows for production.
ต้นฉบับ: The eternal complement
เรียงความ OpenAI โต้แย้งว่า AI ขั้นสูงอาจมีความสําคัญมากที่สุด โดยการเร่งการทํางานการประหารชีวิตตามปกติ หลังการพัฒนา ซึ่งอาจทําให้เกิดความก้าวหน้าทางเศรษฐกิจ
Customer story: grocery retailer Albertsons Companies describes using ChatGPT Enterprise and the OpenAI API to speed up internal work and make shopping easier for customers.
ต้นฉบับ: The Den frees up 10-15 hours a week to grow with ChatGPT Work
Customer story: social club The Den says ChatGPT Work shortened tasks such as grant applications from days to hours, saving its team 10–15 hours a week.
ต้นฉบับ: Disrupting a coordinated model-distillation campaign
OpenAI บอกมันรบกวนความพยายามที่จะประสานงาน เพื่อสกัดเอาเหตุผลต้นแบบที่มีการป้องกันโดยการกลั่น และเป็น เสริมการป้องกันการจู่โจมดังกล่าว
ต้นฉบับ: Helping small businesses put AI to work
OpenAI เป็นหุ้นส่วนกับ SBDC ของอเมริกา เพื่อขยายการอบรม AI และการสนับสนุนท้องถิ่นสําหรับธุรกิจขนาดเล็ก ร่วมกับรายงานว่าทีมที่ใช้ AI น้อยแค่ไหน
ServiceNow AI describes AutoSynthData, a method for turning an enterprise agent's observed weaknesses into many new, verifiable training tasks that fit the target environment.
Hugging Face introduces the Open TTS Leaderboard to make evaluation of multilingual text-to-speech and voice-cloning models more standardized than today's fragmented, arena-based comparisons.
NVIDIA presents Kumo Tabular, a tabular foundation model that predicts new rows through in-context learning instead of training a separate gradient-boosted model for each task.
ต้นฉบับ: Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents
Multiverse Computing explains its ProvenanceGuard paper, which checks not only whether an MCP agent's claim is true but whether it is attributed to the right source.
ต้นฉบับ: Holo4: powering generalist computer-use agents
H Company released Holo4, agentic computer-use models in 27B dense and 35B-A3B mixture-of-experts sizes on its H Models API, plus an updated Holotron4 Nano.
Mistral AI กําลังเปิดศูนย์มิวนิกสําหรับฟิสิกส์ AI และการวิจัยอุตสาหกรรม AI ทํางานร่วมกับหุ้นส่วนจากอุตสาหกรรมเยอรมัน
NVIDIA presents a 64GB DGX Spark as a way for developers to build and run increasingly capable open models and AI agents locally.
NVIDIA กล่าวว่า OpenAI ของ GPT-6 แอสทรา อุลตร้าฟลายวิ่งบน NVIDIA แบล็คเวล GPUs รุ่นนี้มีอยู่ใน OpenAI API และมีสิทธิ์ใช้งาน ChatGPT และผู้ใช้โคเด็กซ์
NVIDIA makes its case for AI-factory returns, citing a cost of roughly $60 million per megawatt and arguing operators need a clear ROI before committing capital.
ต้นฉบับ: NVIDIA Opens Applications for 2027–2028 Graduate Fellowships With Awards Up to $60,000
NVIDIA opened applications for its 2027–2028 Graduate Fellowship program, with awards of up to $60,000.
ต้นฉบับ: From Training to Production, NVIDIA and CoreWeave Close the Loop on Agentic AI
NVIDIA และ Coreweave พรรณนาความสัมพันธ์อันยาวนานของคู่เครื่องยนต์ร่วม และ AI ที่มุ่งเน้นครอบคลุมเมฆ AI จากการฝึกอบรมถึงการผลิต
ต้นฉบับ: Open-sourcing AstaBrief, the fast report-generation model in Asta
AI2 ปล่อย AstaBrief โมเดลเปิดรุ่น 8B ที่เขียนรายงานทางวิทยาศาสตร์ที่อ้างถึง มันขับเคลื่อนโหมดเร็วของแอสตา และสามารถดาวน์โหลดได้ เพื่อใช้โครงสร้างพื้นฐานของคุณเอง
Ai2 introduced Olmo-core 3, a redesigned and fully open training stack for scaling mixture-of-experts models toward the trillion-parameter range.
MIT researchers present InstructMesh, a tool that produces easy-to-edit designs of everyday objects so both experts and beginners can repair AI-generated 3D models and fabricate them.
MIT's Transit Lab will build an open-source Public Transit Intelligence Hub, funded with $2.1 million from Google.org, to unify transit monitoring, operations and rider communication.
MIT News reports on an AI system that beats top-ranked human Stratego players and is more efficient than other models; the researchers see uses in strategic decision-making.
ต้นฉบับ: The effects of an “algorithmic monoculture” depend on the details
An MIT study of hiring decisions finds that many firms relying on the same algorithm can, in some situations, benefit job seekers.
ต้นฉบับ: Who we become when we talk to machines
MIT professor Sherry Turkle's new book, "Artificial Intimacy," critiques chatbots and the antisocial dynamics she argues they encourage.
เรื่องที่มีชื่อเกี่ยวกับ AI เรียงตามคะแนนในช่วง 36 ชั่วโมงที่ผ่านมา
ต้นฉบับ: Vote on which of Hacker News' challenges for AI have been met
ต้นฉบับ: Identity Management for Agentic AI [pdf] (2025)
ต้นฉบับ: Don't be fooled–LLMs don't reason
คัดเลือกบทความใหม่โดยอัตโนมัติ (ให้ความสำคัญกับบทความที่อยู่ในทั้ง cs.AI, cs.CL และ cs.LG) ไม่ใช่การจัดอันดับคุณภาพ
ต้นฉบับ: Comedic Fool's Gold: Reward Exploits and Countermeasures in Conversational Humor
เราตรวจสอบรางวัลอัตโนมัติ สําหรับการฝึกนางแบบภาษาด้วยอารมณ์ขันสนทนา เพ่ง เล็ง ที่ รางวัล ที่ ได้ จาก การ แสวง ประโยชน์ และ การ แก้ แค้น.
ต้นฉบับ: Robust Is Salient: An Informed Adversary Moves the Optimal Signal onto the Salience Pole
เมื่อศัตรูที่มีข้อมูลร่วมกันของผู้ชมของช่องทางที่บังคับ สัญญาณ ที่ ปก ป้อง ความ จริง ได้ ดี ที่ สุด คือ สัญญาณ ที่ ดี ที่ สุด.
Structured tool calls often fail after only a small number of fields violate a schema or an execution contract.
Automated harness optimization can substantially improve LLM agents by iteratively updating their prompts, tool interfaces, and control logic from execution feedback.
ต้นฉบับ: Generation Provenance Before Behavior Attribution: Auditing Synthetic Speech Research Objects
การ ควบคุม พฤติกรรม เพื่อ สังเคราะห์ ข้อมูล การ ฝึก อบรม เรียก ร้อง ให้ รู้ ว่า อะไร ได้ ผลิต ของ ฝึก แต่ ละ อย่าง ก่อน จะ ประเมิน ว่า ของ นั้น เกิด อะไร ขึ้น.
รายงาน การ บีบ บังคับ สรุป ว่า แบบ จําลอง การ ใช้ ภาษา ที่ ถูก อัด แน่น ได้ รับ การ ย้าย จาก แบบ ที่ หนา แน่น ปกติโดยไดเวอร์เจนซ์ของ KL การใช้งานที่ขึ้นอยู่กับผลลัพธ์ที่หนาแน่นของแบบจําลอง จําเป็นต้องรู้จํานวนของการตัดสินใจที่มีการเปลี่ยนแปลง
We propose ReHoPER, an inference-only, zero-shot method that improves large language models' reasoning by generating and answering intermediate questions along multiple paths before the final answer.
การ วัด ความ รู้ มี จุด มุ่ง หมาย ที่ จะ ถ่ายทอด ความ รู้ ที่ แท้ จริง เกี่ยว กับ แบบ จําลอง ภาษา ขนาด ใหญ่ ไป ยัง แบบ จําลอง ขนาด เล็ก เพื่อ ใช้ อย่าง มี ประสิทธิภาพ.
การเรียนรู้แบบข้อความ (ICL) สามารถทําให้โมเดลภาษา สามารถทํางานใหม่ๆ ได้จากการสาธิตโดยไม่เพิ่มน้ําหนัก
This work presents Diffusion Layer Integrated Gradients (DLIG), a token attribution method for diffusion language models (DLMs) that extends Integrated Gradients (IG~\cite{sundararajan2017axiomatic}) to arbitrary layers and denoising steps.
Replay selectors often rank cached trajectories by format feedback, confidence, freshness, or response length, although cache-level correctness and downstream learner utility are distinct objectives.
การแทรกแซงก่อนฝึก เป็นสิ่งสําคัญในการจัดลําดับการวิจัย ตั้งแต่ความเชื่อก่อตัวขึ้นในช่วงก่อนฝึก การจําลองแบบทั่วไปจากการฝึกภายหลัง
Recent years have witnessed the rapid adoption of reinforcement learning (RL) in large language model (LLM) post-training, with substantial gains in mathematical reasoning and code generation.
Introducing new capabilities to frontier models has long been the goal of posttraining, which predominantly employs supervised finetuning (SFT) and reinforcement learning (RL) to this end.
Frequency-collapse attention [Zeris, 2026e] achieves large gains over standard dot-product attention by replacing the Q/K dot product with a bandpass-filtered inner product at a learned frequency.
ไม่ได้รวม:
อัปเดต: 2026-10-02 16:45 UTC