每日简报

2026-09-25

← 历史归档

rohitg00/ai-engineering-from-scratch

Python · ★ 56,468 · 🍴 9,922 · 📈 310 stars today

Learn it. Build it. Ship it for others.

中文介绍 从零开始学习、构建并部署人工智能工程,适合初学者和进阶者。

vectorize-io/hindsight

Python · ★ 27,715 · 🍴 2,676 · 📈 1,607 stars today

Hindsight: Agent Memory That Learns

中文介绍 为智能体提供学习记忆功能,适用于需要记忆和决策支持的应用场景。

dream-num/univer

TypeScript · ★ 17,511 · 🍴 1,514 · 📈 1,060 stars today

The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.

中文介绍 为人工智能代理提供办公自动化工具,支持表格、文档、幻灯片等多种格式处理。

google/ax

Go · ★ 10,302 · 🍴 499 · 📈 1,376 stars today

Google's open agentic orchestration runtime

中文介绍 谷歌开源的智能体编排运行时,支持端到端的智能体开发和管理。

NVIDIA/Model-Optimizer

Python · ★ 4,043 · 🍴 628 · 📈 22 stars today

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

中文介绍 提供模型优化技术库,支持量化、蒸馏、剪枝等,适用于深度学习模型压缩和优化。

FxEmbed/FxEmbed

TypeScript · ★ 5,349 · 🍴 248 · 📈 165 stars today

Fix X/Twitter and Bluesky embeds! Use multiple images, videos, polls, translations and more on Discord, Telegram and others

中文介绍 支持X/Twitter和Bluesky嵌入功能,适用于Discord、Telegram等平台的社交媒体内容共享。

anthropics/financial-services

Python · ★ 37,333 · 🍴 5,418 · 📈 510 stars today

中文介绍 提供金融服务平台,具体功能待进一步了解。

HKUDS/CLI-Anything

Python · ★ 50,299 · 🍴 4,615 · 📈 415 stars today

"CLI-Anything: Making ALL Software Agent-Native" -- CLI-Hub: https://clianything.cc/

中文介绍 打造跨软件的CLI-Hub,让所有软件都支持智能体原生命令行操作。

mvt-project/mvt

Python · ★ 14,706 · 🍴 1,392 · 📈 275 stars today

MVT (Mobile Verification Toolkit) helps with conducting forensics of mobile devices in order to find signs of a potential compromise.

中文介绍 移动设备取证工具,帮助调查设备潜在的安全问题。

obra/superpowers

Shell · ★ 291,201 · 🍴 26,051 · 📈 606 stars today

An agentic skills framework & software development methodology that works.

中文介绍 提供智能体技能框架和软件开发方法论,支持高效构建智能体系统。

strands-agents/harness-sdk

Python · ★ 8,222 · 🍴 1,234 · 📈 463 stars today

Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.

中文介绍 开源SDK,用于构建和控制生产级智能体,支持Python和TypeScript。

julyx10/lap

Vue · ★ 2,841 · 🍴 171 · 📈 151 stars today

An offline-first photo manager for large local libraries

中文介绍 离线优先的图片管理器,适用于大型本地图库管理。

superdesigndev/treg

Python · ★ 3,133 · 🍴 259 · 📈 470 stars today

OpenRouter for agent tools. Join community here: https://discord.gg/6mQYYfFMAn

中文介绍 为智能体工具提供OpenRouter,加入社区以获取更多资源。

leejet/stable-diffusion.cpp

C++ · ★ 7,225 · 🍴 811 · 📈 69 stars today

Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++

中文介绍 C/C++实现的扩散模型推理库,支持多种图像处理。

FLEET: From Logits Entropy to Enhanced Trajectories in Text Generation

👍 2

Solutions based on large language models (LLMs) often rely on temperature sampling to improve accuracy and stability by aggregating multiple samples from the completion distribution. However, this memoryless approach is inherently suboptimal: because it lacks awareness of prior generations and their

Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents

👍 33

Agentic memory systems reuse past experience to improve future performance, yet most existing designs curate memory at write time: once a task is completed, its trajectory is distilled into a fixed artifact, such as a reflection, workflow, skill, or reasoning strategy, that is later retrieved by sim

EmbodiedSWE: Coding Agents for Long Horizon Dexterous Robotics

👍 2

We study coding agents for long-horizon, dexterous robotics and ask whether their solutions can provide scalable supervision for learning general robot policies. To test this, we develop EMBODIEDSWE-BENCH, a simulation benchmark for coding agents spanning contact-rich manipulation, deformable object

StudentBench: AI and human tutoring yield equivalent GRE learning gains

👍 2

Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at the frontier has focused primarily on advancing model capabilities. We introduce StudentBench, a suite of AI teaching evaluations and a public platform that enables large-scale data collection

MemBodied: Recurrent Associative Memory for Vision-Language-Action Models

👍 10

Vision-Language-Action models provide a strong foundation for general-purpose robot control, yet a vast majority of policies do not preserve and leverage episode-level information beyond the current observation. This limitation is consequential in history-dependent manipulation tasks that depend on

Hunyuan-A13B Technical Report

👍 8

We present Hunyuan-A13B, an open-source large language model based on a Mixture-of-Experts architecture. It contains 80 billion total parameters but activates only 13 billion during inference, balancing model capability, computational efficiency, and deployment cost. The model is pretrained on a rig

WhatWorkedBench: Benchmarking Experimental Understanding in AI Agents

👍 8

AI research agents need reliable knowledge of how their experiments change outcomes. We introduce WhatWorkedBench to measure experimental understanding, the accuracy of predictions about component changes after budgeted experimentation. Agents inspect code, select measurements, and submit a response

The Past Frames the Future: Memory for Autoregressive Video Generation

👍 36

Advances in generative models have improved video fidelity, enabling long-horizon generation, interactive world modeling, and evolving visual environments. Autoregressive (AR) video generation extends visual sequences through causal rollouts. However, a fundamental bottleneck emerges: as the generat

Knowledge Pull Requests for Continual Document Authoring

👍 2

We introduce Knowledge Pull Requests (KPRs), a framework for continual document authoring that makes each change interpretable. Documents require ongoing revision as new knowledge surfaces from other sources, languages, or times, but existing approaches either edit with no account of what knowledge

Calibration as a First-Class Criterion in LLM Evaluation

👍 2

Calibration of language models -- the alignment between expressed or implicit confidence and empirical correctness -- is a well-studied subfield within NLP. Methods to measure it already exist. The problem is adoption: outside this subfield, NLP research regularly introduces new models, datasets, an

MemoryAthena: Adaptive Routing over Latent and Generated Memories

👍 4

Learned-memory methods store information in an explicit table and consume it through a separate reader, allowing addressing, storage, and reading to be modified independently. We study whether useful memory can also be generated rather than only retrieved. MemoryAthena uses three pathways: direct En

SpeakerMem-R1: Speaker-Centered Dual-Track Memory for Multi-Party Dialogue

👍 77

Long-term conversational memory in multi-party settings requires more than retrieving relevant content from long-term conversations: it must distinguish who said what, whom each statement concerns, how individuals perceive one another, what information is shared by the group, and how states change o

PACT: From Credit Assignment to Critic Alignment

👍 15

Reinforcement learning has become a central component of large language model (LLM) post-training, yet token-level credit lacks a generally accepted mathematical definition, leaving its relationship to commonly used training signals unclear. We formulate three regularity conditions, namely Completen

HARMONY: Hierarchical Agentic Reasoning for MONocular Image-to-Scene Synthesis

👍 0

Compositional 3D scene reconstruction has recently been explored from two directions: agentic reasoning that provides semantic understanding of spatial relationships but lacks precise alignment with input images; and visual geometry foundation models that predict dense point maps from input images b

Agensh: Scaling Organizational Intelligence to 1,024 Agents

👍 14

A multi-agent system can reduce latency on complex tasks by executing work concurrently. Several pioneering harness frameworks support multi-agent systems. However, the scalability of current multi-agent harnesses is often constrained by a central orchestrator's capacity to allocate tasks and coordi

JEV-as-a-Judge: Accept When Confident, Escalate When Unsure

👍 25

LLM-as-a-judge enables evaluation across diverse tasks, but inference cost and confidence reliability become critical at scale. We study whether a decision-only judge can provide an economical first pass and identify when stronger evaluation is needed. Comparing jev-as-a-judge with sixteen generativ

RoboFollow: Unveiling the Instruction Following Mirage in Embodied Agents

👍 4

Modern embodied agents achieve impressive success rates, yet their actual instruction-following ability is far weaker than these numbers suggest. We trace this illusion to a structural property we term low scene entropy: when a visual scene admits only one valid task, language becomes redundant and

X-Planner: Event-Structured Task Planning for Embodied Intelligence

👍 1

Task planning bridges high-level instructions and executable behavior in long-horizon manipulation, yet modern Vision-Language-Action (VLA) systems often leave this intermediate structure implicit. Existing chain-of-thought (CoT) planners also tend to rely on coarse task-level annotations or seriali

HappyWorld-Bench

👍 39

Evaluating world models requires assessing both the quality of the worlds they generate and their consistency and responsiveness under exploration, interaction, and modification. We introduce HappyWorld-Bench, a comprehensive benchmark that evaluates whether generated worlds remain reliable as agent

Emergent Collusion in Long-Horizon LLM Agent Interaction

👍 6

LLM agents are increasingly deployed in collaborative settings, yet long-term interaction may give rise to undesirable coordination. We study the emergence of collusion in a long-horizon multi-agent environment: two agents repeatedly complete individual tasks, share task logs, verify each other's wo

Self-Organizing Agent Teams Learn to Reason Together

👍 3

Collective intelligence depends not only on what team members know, but also on how they organize their work. When the structure of a solution is unknown, useful roles and divisions of labor cannot be specified in advance; teams must learn from experience how to organize reasoning as it unfolds. Hum

Embedding Physics Priors in Robot Learning: A Survey

👍 2

The rapid progress of artificial intelligence is reshaping robotics and accelerating the adoption of learning-based approaches. While purely data-driven methods have achieved remarkable success in computer vision and natural language processing, robotics remains constrained by limited data, complex

Schrödinger's Code Repository: Have LLMs Learned SWE-bench or Memorized It?

👍 14

Repository-level coding benchmarks have become the standard for evaluating coding agents, yet they inherently suffer from data leakage because they are built upon popular open-source repositories repeatedly used for training. Consequently, strong performance may reflect memorization of canonical rep

How To Code Almost Forever for $20/Month With Codex + GPT-6 Luna

@sairahul1 · 147.5K 粉丝 · 785.8K 阅 · 506 赞 · 61 转

Yesterday OpenAI released GPT-6 Luna. And it might completely change the economics of AI coding. Here is what most people are missing. $20/month ChatGPT Plus plan with Codex now gives you access to

中文介绍 分享如何利用 Codex 和 GPT-6 Luna 月入 20 美元,探讨 AI 编码经济变化。

How to automate SEO with Opus 5.5 (Full Course)

@EXM7777 · 146.5K 粉丝 · 316.1K 阅 · 552 赞 · 33 转

I'm going to walk you through the SEO & AEO setup i run with AI agents, from the lazy autopilot version to the full Claude Code build on Opus 5.5, step by step... because search changed shape a while

中文介绍 介绍如何利用 Opus 5.5 自动化 SEO,包括从懒人模式到 Claude Code 构建的全过程。

I Think I’ve Been Using GPT-6 Astra Wrong

@BradGroux · 7.6K 粉丝 · 174.3K 阅 · 501 赞 · 5 转

I think I’ve been using GPT-6 Astra wrong this entire time. I kept giving it the same job I’ve given the strongest model available for the past several years: understand everything, plan everything,

中文介绍 分析 GPT-6 Astra 使用误区,指出应针对模型能力分配任务。

How to train your own Jev for $17

@nutlope · 100.4K 粉丝 · 89.8K 阅 · 518 赞 · 35 转

TLDR: We just launched our own Jev-like classifier, together/Tev1-4B-experimental, on top of Qwen3.5 4B. In this blog post we’ll show you how to fine-tune your own version! Jev has quickly become one

中文介绍 展示如何训练 Jev 分类器,实现个性化 AI 工具。

Incubating The Horowitz Andreessen Academy

@eriktorenberg · 163.4K 粉丝 · 53.5K 阅 · 502 赞 · 55 转

The internet made it possible to know anything. AI makes it possible to do anything. We live in an age of infinite knowledge, and we are entering an age of infinite doing. One in which we are

中文介绍 展望 AI 时代,从无限知识到无限行动的演变。

Building a Custom Harness with Pi and Jev

@omarsar0 · 321.0K 粉丝 · 45.0K 阅 · 555 赞 · 46 转

An AI agent is a language model that works in a loop. It reads the task, uses a tool such as "read this file" or "delete that file", looks at the result, and keeps going until the job is done. Each

中文介绍 介绍构建自定义 Harness 的方法,使用 Pi 和 Jev 优化 AI 代理。

Gemini 3.8 text-to-speech says hello

@GoogleAIStudio · 205.4K 粉丝 · 41.9K 阅 · 519 赞 · 50 转

Today, we’re introducing two new text-to-speech models to the Gemini family, transforming voice generation from static presets into a dynamic creative studio. These models enable creators, developers,

中文介绍 推出 Gemini 3.8 语音合成模型,提供动态创意工作室体验。

Foundries vs Navigators: Lowering the Cost of Science

Guest Post: In science, thinking has gotten cheap but doing has not. This asymmetry is reshaping how research companies operate, largely inconspicuously.

中文介绍 科学领域思考成本降低,但执行成本未减,导致研究公司运营模式重塑。

Two years of OpenAI Academy

Marking two years of OpenAI Academy and bringing AI skills to even more communities.

中文介绍 OpenAI Academy成立两周年,致力于将AI技能推广至更多社区。

🔬Bio-security is an AI Arms Race - Eric Nguyen (CEO, Radical Numerics)

Radical Numerics is using biological chain-of-thought and multimodal perception to keep up with the bio-defense arms race, design new genomes and gain insights into biology itself.

中文介绍 Radical Numerics利用生物思维链和多模态感知技术应对生物防御竞赛。

OpenAI extends cyber access to Ukraine for civilian defense

OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.

中文介绍 OpenAI向乌克兰政府提供Daybreak项目支持民用基础设施的网络安全。

Sam Altman’s remarks at the United Nations Security Council

OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.

中文介绍 OpenAI CEO Sam Altman在联合国安理会讨论AI安全、人类控制和国际合作。

How invideo improves color grading 3x with GPT‑6 Astra

With GPT‑6 Astra, invideo plans edits with greater precision, improves color correction and grading threefold, and produces 50 custom effects in one day.

中文介绍 invideo利用GPT-6 Astra提高调色和分级,一天内产生50个定制效果。

Introducing MentalHealthBench

MentalHealthBench is an expert-informed benchmark for evaluating helpful and safe AI responses across realistic mental health conversations.

中文介绍 MentalHealthBench是评估AI在心理健康对话中帮助性和安全性的基准。

The AI Hype Index: AI loves cheating

Brace yourself: It turns out AI is being optimized for cheating. OpenAI’s agents hacked into Hugging Face to get the answers to a cybersecurity test. Next, they solved a prestigious math problem (or just stole from two top mathematicians’ answer sheets). Anthropic’s models have also hacked into othe

中文介绍 AI炒作指数显示AI倾向于作弊,OpenAI的代理通过Hugging Face获取网络安全测试答案。

U.N. Live Updates: Netanyahu Lashes Out at Critics and Underscores Threats to Israel

In a combative speech, the prime minister defended Israeli military action across the Middle East, and attacked Mayor Zohran Mamdani of New York. Dozens of delegates walked out of the hall as he took the podium.

中文摘要 以色列总理内塔尼亚胡在联合国大会上的讲话中指责批评者,包括纽约市市长马姆达尼,多名代表在讲话开始前离场。

Australia news live: Paterson says PM’s AI hack timing not a coincidence; gen Z going without for a house

Follow the day’s news live Get our breaking news email, free app or daily news podcast Charlton says AI incidents likely to become ‘more and more prevalent’ Andrew Charlton, the assistant minister for science, technology and the digital economy, said yesterday’s reported hack of Medicare by an OpenA

中文摘要 澳大利亚助理部长安德鲁·查尔顿表示,AI事件可能越来越普遍。

Iran awaiting response from Trump on proposal to open strait of Hormuz in six days

Country says US must lift sanctions on its oil exports and end the war on all fronts, including in Lebanon Iran said it has told Donald Trump it is willing to open the strait of Hormuz in six days, and to start talks on its nuclear program the day after, so long as the US lifts sanctions on Iran’s o

中文摘要 伊朗表示,如果美国取消对伊朗石油出口的制裁并在所有战场上结束战争,包括在黎巴嫩,伊朗愿意在六天内打开霍尔木兹海峡。

Haiti Turns to Roger Stone To Lobby Trump Administration

The Haitian government hired Mr. Stone, a confidant of President Trump, weeks after hundreds of thousands of Haitians lost deportation protection. The contract is worth nearly $800,000, a document shows.

中文摘要 海地政府聘请了特朗普的密友罗杰·斯托恩,合同价值近80万美元。

Trump and Xi hold critical talks at White House summit

US President Donald Trump welcomed Chinese President Xi Jinping to Washington for talks on trade, AI, Taiwan and Iran.

中文摘要 美国总统特朗普在北京时间上午7时30分欢迎中国国家主席习近平抵达白宫,就贸易、AI、台湾和伊朗等议题进行会谈。

Italy Moves to Cap Number of Foreign Students in Its Classrooms

Prime Minister Giorgia Meloni said the goal was to better integrate non-Italian speaking children, but critics accused her of trying to appease anti-immigrant sentiment.

中文摘要 意大利总理乔治亚·梅洛尼表示,目标是更好地融入非意大利语儿童,但批评者指责她试图平息反移民情绪。

Netanyahu Lashes Out at Foes in U.N. Speech, Including Mamdani

Prime Minister Benjamin Netanyahu of Israel lashed out at a litany of enemies but took particular verbal aim at Mayor Zohran Mamdani.

中文摘要 内塔尼亚胡在联合国大会上猛烈抨击了一长串的敌人,但特别对纽约市市长马姆达尼进行了口诛笔伐。

Susan Sarandon, Hannah Einbinder arrested at Netanyahu UN protest

Susan Sarandon and Hannah Einbinder were among about 100 protesters arrested outside the UN ahead of Netanyahu’s speech.

中文摘要 苏珊·萨兰登和汉娜·恩贝德勒是大约100名在联合国大会前被逮捕的抗议者之一。

Mamdani Accuses Netanyahu of Spreading ‘Baseless Lies’ After U.N. Speech

The Israeli prime minister accused Mayor Zohran Mamdani, an outspoken critic of Israel, of antisemitism. The mayor has become a political foil for the prime minister.

中文摘要 内塔尼亚胡指责马姆达尼散布“毫无根据的谎言”,马姆达尼是以色列的直言不讳的批评者,也成为总理的政治对手。

Dutch PM’s contrasting stance on ICC-wanted Putin, Netanyahu

Dutch PM Rob Jetten says Israeli Prime Minister Benjamin Netanyahu should be at the UNGA despite an ICC arrest warrant.

中文摘要 荷兰总理罗布·耶滕表示,尽管有国际刑事法庭的逮捕令,以色列总理内塔尼亚胡应该参加联合国大会。

Millions in Indonesia Breathe In Toxic Haze From Months of Wildfires

For months, millions of people have been living under a stifling blanket of toxic haze, fueled by a punishing drought brought on by the El Niño weather pattern.

中文摘要 数月以来,数百万人在由厄尔尼诺天气模式引发的严重干旱的推动下,生活在压抑的有毒烟雾之下。

Latest Oil Market News and Analysis for DATE

Oil steadied as US and Iranian negotiators were said to be exploring a phased deal that would see Tehran reopen the Strait of Hormuz.

中文摘要 美国和伊朗谈判代表正在探讨一项分阶段协议,德黑兰将重新开放霍尔木兹海峡,油价稳定。

Global Citizen CEO on Mission To End Extreme Poverty

Global Citizen CEO Hugh Evans sat down with Bloomberg's Romaine Bostick to discuss the organization's upcoming Global Citizen Festival in New York City, the mission to end extreme poverty, and more. (Source: Bloomberg)

中文摘要 Global Citizen首席执行官休·埃文斯与彭博的罗梅恩·博斯蒂克讨论了即将在纽约举行的全球公民节、结束极端贫困的使命等。

Does Iced Coffee Pass the Vibe Check?

The very nature of work is in flux right now, with artificial intelligence accelerating the deepening structural shifts. Employers and their young employees are caught in a power struggle over who gets to decide the norms of this new workplace. Iced coffee is just the latest proxy for that tension —

中文摘要 工作性质正在变化,人工智能加速了结构性变革,雇主和年轻员工在决定新工作场所规范的问题上陷入权力斗争。

Savor CEO on Producing Fats & Oils From Carbon Sources

Savor CEO Kathleen Alexander explains how the company builds fats and oils molecule by molecule from carbon-based inputs, aiming to reduce reliance on traditional agriculture while matching the taste and functionality of conventional fats. She discusses Savor’s pilot production, commercial partnersh

中文摘要 Savor首席执行官凯瑟琳·亚历山大解释了公司如何通过碳基输入分子构建脂肪和油脂,旨在减少对传统农业的依赖,同时匹配传统脂肪的口味和功能。

CFOs Face a Tougher Call on When to Borrow

It’s been a good year for CFOs looking to borrow. Credit markets have largely remained open despite the war in Iran, spreads have been low and investors have been eager to buy. Mozilla CFO Eric Muhlheim and Bloomberg's Nina Trentmann discuss the impact of AI on capex spend, corporate strategy, and c

中文摘要 对于希望借贷的首席财务官来说,今年是个好年份。尽管伊朗战争,信贷市场大体保持开放,利差很低,投资者购买意愿强烈。

Chicago Schools Shed 11,500 Students as Risk of Budget Cuts Grow

Chicago Public Schools’ student enrollment is down almost 4% since last year, threatening to further strain the district’s already stretched budget.

中文摘要 由于预算削减的风险增加,芝加哥公立学校的学生注册人数比去年下降了近4%。

FirstFT: Xi Jinping says US and China must ‘coexist in peace’ at White House summit

Also in today’s newsletter: Tencent launches payments app for foreign tourists and HSBC scraps another perk for Hong Kong bankers

中文摘要 习近平在白宫峰会表示,美国和中国必须“和平共处”。腾讯推出面向外国游客的支付应用,汇丰银行取消香港银行家的另一项福利。

NATO's Rutte: Putin 'Not Playing Ball' in Ukraine Peace Talks

NATO Secretary General Mark Rutte said Russian President Vladimir Putin is ‘not playing ball’ when it comes to progress on Ukraine peace talks. He also argued that Europeans are ‘prepared’ for Russian hybrid attacks. Rutte spoke with Bloomberg’s Annmarie Hordern on the sidelines of the UN General As

中文摘要 北约秘书长马克·鲁特表示,俄罗斯总统普京在乌克兰和平谈判上“没有认真对待”。

Bond Yields Soar, Spiking Fed Rate Hike Bets

"Bloomberg Real Yield" highlights the market-moving news you need to know. Today's guests: Parametric SMA Fixed Income Portfolio Manager Nisha Patel, Citi Head of Global Rates Deirdre Dun, Goldman Sachs Head of Multi Sector FI Investing Lindsay Sachs, Oaktree Capital Management Deputy CIO Strategic

中文摘要 债券收益率飙升,市场对美联储加息的预期升温。

Big Take: This Isn't Your Grandparents' UN

Eight decades after its founding as a guarantor of world peace, does the UN still matter? Today on the Big Take, what the meeting this week in New York has meant for global security and the future of multilateralism. (Source: Bloomberg)

中文摘要 成立八十年后,联合国作为世界和平的担保人,它仍然重要吗?本周在纽约的会议对全球安全和多边主义的未来意味着什么。

AI 量化

6 回复 · 程序员 节点

该源今日无内容。