每日简报

2026-09-27

← 历史归档

paperclipai/paperclip

TypeScript · ★ 87,185 · 🍴 15,418 · 📈 2,589 stars today

The open-source app everyone uses to manage agents at work

中文介绍 Paperclip 是一款开源应用,用于管理工作中的智能代理,简化工作流程。

vectorize-io/hindsight

Python · ★ 32,100 · 🍴 3,565 · 📈 2,152 stars today

Hindsight: Agent Memory That Learns

中文介绍 Hindsight 是一款智能代理记忆学习工具,用于构建智能代理的长期记忆。

NVIDIA/Model-Optimizer

Python · ★ 4,729 · 🍴 667 · 📈 354 stars today

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

中文介绍 NVIDIA Model Optimizer 提供一系列先进的模型优化技术,用于压缩深度学习模型,提高效率。

dream-num/univer

TypeScript · ★ 19,182 · 🍴 1,633 · 📈 845 stars today

The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.

中文介绍 Univer 是一款办公套件,集成表格、文档、幻灯片、画布、关系表和PDF等功能,适用于智能代理。

tensorflow/tensorflow

C++ · ★ 200,438 · 🍴 77,606 · 📈 31 stars today

An Open Source Machine Learning Framework for Everyone

中文介绍 TensorFlow 是一个开源的机器学习框架,适用于广泛的数据分析和机器学习任务。

rohitg00/ai-engineering-from-scratch

Python · ★ 58,333 · 🍴 10,117 · 📈 828 stars today

Learn it. Build it. Ship it for others.

中文介绍 AI Engineering From Scratch 是一个学习资源,帮助初学者从零开始学习人工智能工程。

openbao/openbao

Go · ★ 7,989 · 🍴 590 · 📈 360 stars today

OpenBao is a software solution to manage, store, and distribute sensitive data including secrets, certificates, and keys.

中文介绍 OpenBao 是一款软件解决方案,用于管理、存储和分发敏感数据,如机密、证书和密钥。

block/buzz

Rust · ★ 34,816 · 🍴 4,594 · 📈 367 stars today

A hive mind communication platform

中文介绍 Buzz 是一个蜂群思维沟通平台,用于集体智慧和协作。

microsoft/vscode

TypeScript · ★ 193,062 · 🍴 43,534 · 📈 78 stars today

Visual Studio Code

中文介绍 Visual Studio Code 是一款流行的代码编辑器,支持多种编程语言和开发环境。

zhaoxuya520/reverse-skill

PowerShell · ★ 37,973 · 🍴 5,268 · 📈 409 stars today

Reverse Engineering / Authorized Penetration Testing / Security Research Skill Router Pack AI-powered routing + On-demand toolchain bootstrapping + Self-evolving knowledge base Supports Claude Code, Kiro, Cursor, Cline, and other AI coding clients 逆向/渗透/安全技能路由包 - AI 自动路由 + 按需自举工具链 + 自动进化经验库 | 支持 Cla

中文介绍 Reverse Skill 是一个逆向工程、授权渗透测试和安全研究技能路由包,支持 Claude Code 等工具。

llvm/llvm-project

LLVM · ★ 40,743 · 🍴 18,823 · 📈 29 stars today

The LLVM Project is a collection of modular and reusable compiler and toolchain technologies.

中文介绍 LLVM 项目是一个模块化和可重用的编译器和工具链技术集合。

anthropics/claude-code-action

TypeScript · ★ 9,076 · 🍴 2,162 · 📈 15 stars today

中文介绍 Claude Code Action 是一个用于代码操作的工具。

actions/runner-images

PowerShell · ★ 13,290 · 🍴 3,864 · 📈 13 stars today

GitHub Actions runner images

中文介绍 GitHub Actions runner images 提供了 GitHub Actions 运行器所需的镜像。

mobile-next/mobile-mcp

TypeScript · ★ 7,318 · 🍴 640 · 📈 143 stars today

Model Context Protocol Server for Mobile Automation and Scraping (iOS, Android, Emulators, Simulators and Real Devices)

中文介绍 Mobile MCP 是一款移动自动化和抓取的模型上下文协议服务器,支持iOS、Android、仿真器和真实设备。

vercel/next.js

JavaScript · ★ 142,614 · 🍴 33,146 · 📈 31 stars today

The React Framework

中文介绍 Next.js 是一个React框架,用于构建服务器端渲染的Web应用。

RGBD20K: A Large-Scale Benchmark for RGB-D Semantic Segmentation

👍 7

In this paper, we propose RGBD20K, a novel dataset for facilitating the development of more robust and general RGB-D semantic segmentation by encompassing abundant categories and high-quality annotations. RGBD20K possesses several attractive properties: (1) Expanded Semantic Space. In particular, it

Coding Agents for Generalized Task and Motion Planning Problems

👍 9

Task and motion planning (TAMP) problems remain difficult even with full observability and object-centric states because discrete decisions are tightly coupled to geometric, kinematic, and dynamic constraints. Generalized TAMP addresses this difficulty by exploiting regularities across problem insta

Rufus-Air: An Open LLM Post-Training Recipe

👍 13

Rufus-Air is an open and reproducible post-training recipe on GLM-4.5-Air-Base (106B-A12B), organized as a serial pipeline of eight stages: SFT, Reasoning RL, Coding RL, Instruction-Following RL, General Agent, Coding Agent, Search Agent, and RLHF. We document the data, reward design, infrastructure

IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis

👍 9

Deep search requires LLM agents to decompose complex queries, search for evidence, and synthesize grounded answers, yet existing ReAct-style agents suffer from two limitations: role coupling, where one policy must handle planning, evidence use, and synthesis; and context accumulation, where growing

PUBG Ally: A Conversational Embodied Agent as an AI Teammate

👍 5

We introduce PUBG Ally, an embodied agent for PUBG: BATTLEGROUNDS that can reason, act autonomously, and play alongside players as a voice-enabled teammate. Building such a teammate requires combining two difficult capabilities: it must perceive and respond to a constantly changing game world under

Learning to Discover Interesting Mathematics

👍 8

Recently, Large Language Models (LLMs) have been increasingly able to solve advanced mathematical problems, including many that have been open for decades. This opens the door to expansion of mathematical knowledge at unprecedented scale. Yet, while LLMs may be able to conjecture and prove more and

DeltaWAM: Delta World Action Models for Bimanual Manipulation

👍 3

World-action models (WAMs) transfer visual and motion priors from pretrained video generators to robot control by jointly modeling visual dynamics and actions. Existing WAMs, however, predict dense future frames during training, repeatedly modeling largely unchanged content and coupling action-condi

Training Object Permanence in World Models

👍 197

Object permanence and solidity are hallmarks of human cognitive priors. Recent studies show that video generation models, a paradigmatic class of current world models, have begun to show emerged reasoning abilities, making them ideal candidates for building human-like physical intelligence. Do video

Agent-Editing World Model: Rethinking World Modeling for LLM Agents

👍 17

Recent advances in large language models (LLMs) have enabled agents to tackle long-horizon tasks across diverse environments. To further improve agent performance, existing language world models typically predict environment observations, yet reconstructing high-entropy, execution-dependent tool res

FLEET: From Logits Entropy to Enhanced Trajectories in Text Generation

👍 4

Solutions based on large language models (LLMs) often rely on temperature sampling to improve accuracy and stability by aggregating multiple samples from the completion distribution. However, this memoryless approach is inherently suboptimal: because it lacks awareness of prior generations and their

Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents

👍 34

Agentic memory systems reuse past experience to improve future performance, yet most existing designs curate memory at write time: once a task is completed, its trajectory is distilled into a fixed artifact, such as a reflection, workflow, skill, or reasoning strategy, that is later retrieved by sim

EmbodiedSWE: Coding Agents for Long Horizon Dexterous Robotics

👍 3

We study coding agents for long-horizon, dexterous robotics and ask whether their solutions can provide scalable supervision for learning general robot policies. To test this, we develop EMBODIEDSWE-BENCH, a simulation benchmark for coding agents spanning contact-rich manipulation, deformable object

StudentBench: AI and human tutoring yield equivalent GRE learning gains

👍 3

Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at the frontier has focused primarily on advancing model capabilities. We introduce StudentBench, a suite of AI teaching evaluations and a public platform that enables large-scale data collection

Knowledge Pull Requests for Continual Document Authoring

👍 2

We introduce Knowledge Pull Requests (KPRs), a framework for continual document authoring that makes each change interpretable. Documents require ongoing revision as new knowledge surfaces from other sources, languages, or times, but existing approaches either edit with no account of what knowledge

Calibration as a First-Class Criterion in LLM Evaluation

👍 2

Calibration of language models -- the alignment between expressed or implicit confidence and empirical correctness -- is a well-studied subfield within NLP. Methods to measure it already exist. The problem is adoption: outside this subfield, NLP research regularly introduces new models, datasets, an

MemoryAthena: Adaptive Routing over Latent and Generated Memories

👍 5

Learned-memory methods store information in an explicit table and consume it through a separate reader, allowing addressing, storage, and reading to be modified independently. We study whether useful memory can also be generated rather than only retrieved. MemoryAthena uses three pathways: direct En

X-Planner: Event-Structured Task Planning for Embodied Intelligence

👍 2

Task planning bridges high-level instructions and executable behavior in long-horizon manipulation, yet modern Vision-Language-Action (VLA) systems often leave this intermediate structure implicit. Existing chain-of-thought (CoT) planners also tend to rely on coarse task-level annotations or seriali

HappyWorld-Bench

👍 41

Evaluating world models requires assessing both the quality of the worlds they generate and their consistency and responsiveness under exploration, interaction, and modification. We introduce HappyWorld-Bench, a comprehensive benchmark that evaluates whether generated worlds remain reliable as agent

OmniEcho: Spatial Audio Understanding for Embodied Agents

👍 20

Humans can effortlessly localize the direction of a sound source and integrate it with visual cues for reasoning, yet this remains challenging for embodied agents. In particular, it is still unclear how to effectively evaluate and model spatial audio understanding in embodied settings. To address th

Self-Organizing Agent Teams Learn to Reason Together

👍 5

Collective intelligence depends not only on what team members know, but also on how they organize their work. When the structure of a solution is unknown, useful roles and divisions of labor cannot be specified in advance; teams must learn from experience how to organize reasoning as it unfolds. Hum

AgentKernel: The Trust-Native Agentic Operating System

👍 6

Modern AI agents routinely cross trust boundaries: they ingest untrusted content, combine it with privileged instructions, persist intermediate beliefs in long-term memory, and invoke privileged tools. This creates an attack surface in which malicious payloads can enter through model inputs and caus

The New Copilot

@jacobandreou · 7.8K 粉丝 · 521.0K 阅 · 500 赞 · 77 转

The bottleneck is no longer intelligence. Over the last six months, intelligence has accelerated dramatically. But smarter models are not enough for everyone to benefit from novel intelligence. New

中文介绍 博主讨论了人工智能发展的瓶颈,指出智能模型加速但仍需更多创新以使所有人受益。

We're open sourcing the company brain, Here's how we designed the multi-player harness

@DhravyaShah · 63.2K 粉丝 · 112.4K 阅 · 608 赞 · 45 转

It's been just 2 weeks since we discontinued the company brain harness. We had to offboard tons of people to other products, but many of our customers, and others suggested us to open source our work

中文介绍 公司开源了其多用户协作工具,并分享了设计思路,尽管之前停止了相关服务。

Why AI is booming, but productivity isn't

@chamath · 2.4M 粉丝 · 97.8K 阅 · 551 赞 · 44 转

In June, I wrote that vibe coding was dead and that ROI-driven analysis of AI was about to go from a nice-to-have to a necessity. Here's what AI ROI is and why it's tricky to see in GDP numbers so

中文介绍 博主分析了AI繁荣但生产力未提升的原因,探讨了AI投资回报率在GDP数据中的难以体现。

Prescriptions for Prosperity in the Digital Economy

@saylor · 5.2M 粉丝 · 96.2K 阅 · 661 赞 · 98 转

Artificial intelligence will make it possible for individuals and companies to produce far more than they can today. That makes the freedom to create, finance, own, and exchange things more important.

中文介绍 博主提出了数字经济发展中创造、融资、拥有和交换物品自由的必要性,认为人工智能将极大提升生产效率。

Jev is a state-of-the-art search reranker

@ErikKaum · 2.0K 粉丝 · 75.0K 阅 · 501 赞 · 44 转

We evaluated Jev as a reranker for search results in turbopuffer. The results may surprise you. Jev isn’t just a good reranker, it performed better than existing state-of-the-art rerankers in our

中文介绍 博主评估了Jev作为搜索结果重排器的性能,指出其在某些方面超越了现有最佳重排器。

Using Claude Code: Spending your effort

@trq212 · 354.7K 粉丝 · 62.7K 阅 · 965 赞 · 52 转

One of the best parts of our newest Claude models is how they respond to effort without breaking the prompt cache in Claude Code, but I’ve received a lot of questions on this from users. What is

中文介绍 博主解答了关于Claude模型如何处理用户努力的问题,并讨论了相关功能。

We built a 20x faster Grok bot

@cerebras · 76.6K 粉丝 · 37.6K 阅 · 554 赞 · 28 转

Written by @milksandmatcha and @0xSero A personal assistant should save you time and effort. Over the past few weeks, we’ve been obsessively testing AI personal assistants on everyday tasks, from

中文介绍 介绍了一款速度提升20倍的Grok机器人,强调了个人助理在日常任务中的时间节省功能。

Ghast AI Season 3 Starts September 25

@Ghast_AI · 10.0K 粉丝 · 6.3K 阅 · 542 赞 · 31 转

We just introduced the Capability Economy, a market where private AI work can be bought and verified without exposing the method behind it. Every Capability in that market starts as a workflow someone

中文介绍 Ghast AI推出新功能“能力经济”,允许用户购买和验证私人AI工作流程而不透露方法。

Proaction boosts sales 60% and saves 75+ hours with Codex

With Codex, GPT-Live-1, and GPT-6 Astra, Proaction builds, operates, and sells modern fleet management faster.

中文介绍 OpenAI的Proaction利用Codex、GPT-Live-1和GPT-6 Astra,将现代车队管理构建、运营和销售速度提高60%,节省超过75小时。

The Pentagon wants $30 million to build an AI-powered lie detector

The US government wants to spend $30.3 million over the next five years on an improved form of lie detector, according to a Department of Defense budget request. The program, called Polygraph+ or Polygraph Next, will focus on scoring algorithms that use artificial intelligence and machine learning a

中文介绍 美国国防部请求在未来五年内投资3030万美元,用于改进的测谎仪技术,名为Polygraph+或Polygraph Next。

[AINews] The Future of Latent Space

A quiet day lets us discuss the work behind the scenes - now open for business!

中文介绍 Latent Space开放业务。

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

GWM Worlds 2 uses persistent context and timed actions to steer a world model generating video and audio in real time.

中文介绍 GWM Worlds 2使用持续上下文和定时动作来引导生成视频和音频的世界模型。

Foundries vs Navigators: Lowering the Cost of Science

Guest Post: In science, thinking has gotten cheap but doing has not. This asymmetry is reshaping how research companies operate, largely inconspicuously.

中文介绍 科学中思考变得廉价,但执行却不贵,这种不对称性正在重塑研究公司的运营方式。

Two years of OpenAI Academy

Marking two years of OpenAI Academy and bringing AI skills to even more communities.

中文介绍 OpenAI Academy成立两周年,将AI技能带给更多社区。

🔬Bio-security is an AI Arms Race - Eric Nguyen (CEO, Radical Numerics)

Radical Numerics is using biological chain-of-thought and multimodal perception to keep up with the bio-defense arms race, design new genomes and gain insights into biology itself.

中文介绍 Radical Numerics利用生物思维链和多模态感知参与生物防御军备竞赛,设计新基因组,并深入了解生物学本身。

OpenAI extends cyber access to Ukraine for civilian defense

OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.

中文介绍 OpenAI将其Daybreak项目扩展到乌克兰政府,以支持民用基础设施的网络安全。

Sam Altman’s remarks at the United Nations Security Council

OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.

中文介绍 OpenAI首席执行官Sam Altman在联合国安理会讨论了AI安全、人类控制和国际合作。

Australia news live: Albanese says announcement of Medicare hack ‘purely in national interest’

Follow live updates Get our breaking news email, free app or daily news podcast Good morning Welcome to our Sunday live blog. Flight delays and cancellations caused frustration and heartache for some footy fans racing to make it to Saturday’s AFL grand final before the first bounce. And the PM says

中文摘要 Albanese表示,Medicare系统遭黑客攻击的声明‘完全符合国家利益’。

Evicted 87-Year-Old Woman Urges Spaniards to ‘Fight’ to Avoid Same Fate

María del Carmen Abascal, whose case has become a symbol of Spain’s housing crisis, appeared in a video from a hospital bed after being removed from her apartment.

中文摘要 被驱逐的87岁老妇人在医院病房中呼吁西班牙人‘反抗’以避免相同命运。

Trump rejects Iran deal to reopen Strait of Hormuz in seven days

Iran's foreign minister acknowledges Trump's comments, but adds that Tehran is waiting for the "definitive views" of mediators.

中文摘要 特朗普拒绝伊朗重启霍尔木兹海峡协议,称伊朗外交部长承认其评论,但表示德黑兰正在等待调解者的‘最终看法’。

Australian backlash to datacentres is faux import from US, officials say: ‘We are not the United States’

Pro-datacentre officials argue scale of Australia’s buildout is smaller, more regulated and less damaging than in US Australian politicians and business leaders are arguing that their populace’s growing resentment of AI datacentres is an inauthentic import from the United States. Like the US, Austra

中文摘要 澳大利亚对数据中心的不满是美国的不实进口,官员们表示‘我们不是美国’。

Trump gave Xi Jinping a warm welcome. What did the talks accomplish?

From the National Archives to the state dinner, President Trump and Xi Jinping looked comfortable together in Washington. Two former NPR China correspondents on what the visit did and didn't deliver.

中文摘要 特朗普热烈欢迎习近平,这次会谈达成了什么成果?

Is Congo's Ebola outbreak coming under control?

Four months in, Congo's Ebola outbreak is one of the largest ever recorded. Officials and outside experts can't agree on whether it's finally slowing down.

中文摘要 刚果(金)埃博拉疫情是否正在得到控制?官员和外界专家对此意见不一。

Trump Rejects Iran’s Cease-Fire Proposal to Reopen Strait of Hormuz

The president said the deal was not “acceptable” and demurred on whether he would restart military strikes after the midterm elections.

中文摘要 特朗普拒绝伊朗关于重新开放霍尔木兹海峡的停火提议,称该协议‘不可接受’,在中期选举后推迟了是否重启军事行动的决定。

Bangladesh Targets Up to $1 Billion in Debut Sovereign Bond Sale

Bangladesh is preparing to raise between $500 million and $1 billion through its maiden foreign-currency sovereign bond sale within the next three months, a senior government official said.

中文摘要 孟加拉国计划在未来三个月内通过首次外币主权债券发行筹集5亿至10亿美元。

Bloomberg This Weekend 09/26/2026

The news doesn’t stop when markets close. Hosts David Gura, Christina Ruffini and Lisa Mateo bring clarity, context and a bit of humor to the weekend’s biggest headlines, LIVE from New York. Joined by The New York Times Kyiv Bureau Chief Andrew Kramer, The Netherlands Prime Minister Rob Jetten, Bull

中文摘要 纽约时报基辅分社社长安德鲁·克拉默等嘉宾参与讨论,探讨周末重要新闻。

Pointed! Bloomberg's Weekly News Quiz For Risk-Takers

Pointed offers a strategic twist to the news quiz format, testing not just players’ knowledge of the news but also their confidence in their answers. Join Bloomberg's Alexis Christoforous, Christina Ruffini and David Gura as they play and check out the quiz for yourself at Bloomberg.com (Source: Blo

中文摘要 Bloomberg的Pointed新闻问答节目增加了策略性,测试参与者对新闻的了解和自信。

US and China Seek Common Ground on AI

Bloomberg News Chief North Asia Correspondent Stephen Engle and ABC News State Department reporter Shannon Kingston tell Bloomberg This Weekend that President Donald Trump and Chinese President Xi Jinping’s summit highlighted efforts to establish dialogue on AI while building trust between the two c

中文摘要 美国和中国在人工智能领域寻求共同立场,同时建立对话。

How Syracuse's Semiconductor Gamble May Help Save the Rust Belt

What does it actually take to bring a Rust Belt city back to life? Micron is building one of the largest semiconductor factories in US history in Syracuse, a city that experienced major job loss after companies including GE, Carrier and General Motors left the area. Rob Simpson of CenterState CEO ex

中文摘要 美光公司在锡拉丘兹建造美国历史上最大的半导体工厂,以帮助振兴该地区。

Birding Apps Bring More People Into the Wild

On Bloomberg This Weekend, writer and photographer Alexandra Marvar explains how apps such as Merlin are making birding easier and more accessible, helping draw new people to a hobby that offers an escape into nature. Speaking with hosts David Gura and Christina Ruffini, Marvar also explores whether

中文摘要 观鸟应用程序如Merlin使观鸟更容易、更易于接触,吸引新爱好者。

Mexico Stresses Sovereignty in US Cooperation

Mexico’s foreign minister Roberto Velasco Álvarez tells Bloomberg This Weekend that the country maintains a strong relationship with the US while insisting cooperation on security must respect Mexico’s sovereignty and territorial integrity. Speaking with hosts David Gura and Christina Ruffini, he sa

中文摘要 墨西哥外长罗伯托·韦尔萨科·阿尔瓦雷斯强调,在美墨合作中维护墨西哥的主权和领土完整。

Russia Ukraine War Becomes Economic Battle

New York Times Kyiv bureau chief Andrew Kramer tells Bloomberg This Weekend that Russia and Ukraine are increasingly targeting each other’s economies as the war shifts toward attacks on energy, transportation and export infrastructure. Speaking with hosts David Gura and Christina Ruffini, Kramer say

中文摘要 俄罗斯和乌克兰在战争中对彼此的经济进行攻击,战争转向对能源、交通和出口基础设施的攻击。

该源今日无内容。