每日简报

2026-08-24

← 历史归档

openai/codex

Rust · ★ 115,126 · 🍴 17,556 · 📈 2,729 stars today

Lightweight coding agent that runs in your terminal

中文介绍 轻量级终端代码执行代理,用于实现代码的即时执行和调试。

freestylefly/awesome-gpt-image-2

JavaScript · ★ 12,680 · 🍴 1,427 · 📈 440 stars today

Prompt as Code | GPT-Image2 工业级提示词引擎与模板库,470+ 个案例逆向工程,20+ 套工业级模板,并提炼出Skills,持续更新中

中文介绍 GPT-Image2 工业级提示词引擎与模板库,提供丰富的案例和模板,用于构建图像生成应用。

mattpocock/skills

Shell · ★ 233,809 · 🍴 19,942 · 📈 2,448 stars today

Skills for Real Engineers. Straight from my .agents directory.

中文介绍 为真实工程师设计的技能库,从个人代理目录中提取技能。

basecamp/omarchy

Shell · ★ 29,099 · 🍴 2,959 · 📈 814 stars today

Beautiful, Modern & Opinionated Linux

中文介绍 美观且现代的 Linux 发行版,具有独特的风格。

AprilNEA/OpenLogi

Rust · ★ 14,898 · 🍴 398 · 📈 1,008 stars today

⚡️A native, local-first alternative to Logitech Options+, written in Rust 🦀 — remap buttons, DPI, and SmartShift over HID++. No account, no telemetry.

中文介绍 Logitech Options+ 的本地化替代品,使用 Rust 编写,支持按键重映射、DPI 调整和 SmartShift 功能。

block/buzz

Rust · ★ 30,090 · 🍴 3,825 · 📈 349 stars today

A hive mind communication platform

中文介绍 一个基于蜂群思维的沟通平台,用于集体智慧和协作。

apache/maka

TypeScript · ★ 2,336 · 🍴 269 · 📈 49 stars today

Apache Maka (Incubating) is a local-first AI agent workspace. Model messages, tool calls, tool results, permission decisions, and termination events are recorded as an append-only log.

中文介绍 本地优先的 AI 代理工作空间,记录模型消息、工具调用、结果、权限决策和终止事件。

Alishahryar1/free-claude-code

Python · ★ 47,935 · 🍴 7,887 · 📈 1,040 stars today

Use Claude Code, Codex, Pi, and OpenCode for free (1.3B+ free tokens) from your terminal, app, IDE, or phone like OpenClaw (voice supported + ToS friendly)

中文介绍 从终端、应用、IDE 或手机免费使用 Claude Code、Codex、Pi 和 OpenCode,支持语音输入。

tinyhumansai/openhuman

Rust · ★ 36,726 · 🍴 3,670 · 📈 106 stars today

Your Personal AI super intelligence. A brain that builds a local-first memory of your life, a fantastic orchestrator of agent fleets and workflows, and a deep researcher.

中文介绍 个人 AI 超智能,构建本地优先的生命记忆,协调代理群和流程,进行深入研究。

affaan-m/ECC

JavaScript · ★ 242,544 · 🍴 36,718 · 📈 427 stars today

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

中文介绍 Claude Code、Codex、Opencode、Cursor 等的代理性能优化系统,提供技能、本能、记忆、安全和研究优先的开发。

ruvnet/ruflo

TypeScript · ★ 69,060 · 🍴 8,273 · 📈 134 stars today

🌊 The original agent meta-harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory, self-learning intelligence, RAG integration, and native Claude Code / Codex / Hermes and many more Integrated

中文介绍 智能多玩家蜂群部署、自主工作流协调和对话式 AI 系统构建的原生元代理工具,具有自适应记忆和自学习智能。

VoltAgent/awesome-agent-skills

★ 31,273 · 🍴 3,363 · 📈 223 stars today

A curated collection of 1000+ agent skills from official dev teams and the community, compatible with Claude Code, Codex, Gemini CLI, Cursor, and more.

中文介绍 1000+ 个代理技能集合,兼容 Claude Code、Codex、Gemini CLI、Cursor 等,适用于官方团队和社区。

virgiliojr94/book-to-skill

Python · ★ 24,633 · 🍴 2,577 · 📈 423 stars today

Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.

中文介绍 将技术书籍 PDF 转换为 Claude Code 技能,便于学习和使用。

dani-garcia/vaultwarden

Rust · ★ 65,952 · 🍴 3,126 · 📈 95 stars today

Unofficial Bitwarden compatible server written in Rust, formerly known as bitwarden_rs

中文介绍 非官方的 Bitwarden 兼容服务器,使用 Rust 编写,提供密码管理服务。

anthropics/claude-plugins-community

Python · ★ 923 · 🍴 126 · 📈 257 stars today

Community plugin marketplace for Claude Cowork and Claude Code. Read-only mirror — submit plugins at clau.de/plugin-directory-submission.

中文介绍 Claude Cowork 和 Claude Code 的社区插件市场,提供插件提交和阅读功能。

ripienaar/free-for-dev

HTML · ★ 134,405 · 🍴 14,062 · 📈 593 stars today

A list of SaaS, PaaS and IaaS offerings that have free tiers of interest to devops and infradev

中文介绍 列出对开发人员有价值的免费 SaaS、PaaS 和 IaaS 服务。

Comfy-Org/ComfyUI

Python · ★ 129,367 · 🍴 15,254 · 📈 179 stars today

The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.

中文介绍 功能强大且模块化的扩散模型 GUI、API 和后端,具有图形/节点界面。

NousResearch/hermes-agent

Python · ★ 234,963 · 🍴 47,334 · 📈 519 stars today

The agent that grows with you

中文介绍 随用户成长而发展的智能代理。

FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM Serving

👍 16

Long-context modeling is a pivotal capability for Large Language Models, yet the quadratic complexity of attention remains a critical bottleneck, particularly during the compute-intensive prefilling phase. Our previous work, FlashPrefill, mitigates this cost through instantaneous pattern discovery a

SWE-bench Science: Can Coding Agents Resolve Engineering Tasks in Science?

👍 61

Software increasingly functions as part of the scientific instrument itself, making failures in scientific code capable of compromising not only program behavior but also the evidence underlying scientific conclusions. Yet existing evaluations of coding agents largely emphasize aggregate task succes

Repo0: Design-Driven Zero-to-All Code Generation

👍 17

Large language model agents have made substantial progress in code generation, yet most existing systems assume a predefined repository architecture. This assumption does not hold in zero-to-all code generation, where an agent must construct an entire software project directly from natural-language

MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use

👍 31

Memory has become a key component of large language models, enabling them to retain information and learn from long-term interactions. However, existing memory benchmarks mainly evaluate whether information is correctly extracted, stored, and retrieved, while largely overlooking how retrieved memori

EnvHarness: Awakening Static Worlds for Agent Learning

👍 254

LLM agents learn by interacting with environments, yet these environments are hand-built and static: blind to an agent's weaknesses, and quickly left behind as it improves. While recent environment generation methods attempt to address this, they require domain-specific pipelines, rely on expensive

FACET: Preserving Source Intent and Executable State in Terminal Task Synthesis

👍 114

Training terminal agents requires scalable executable supervision, yet synthesizing high-quality terminal tasks remains challenging. Each task couples an instruction, an initialized environment, a reference solution, and an executable verifier; if these artifacts are generated from inconsistent assu

SkillGate: Training In-Policy Skill Selection in Long-Horizon Agents

👍 6

Agent frameworks increasingly package procedural knowledge as skills: instruction files an agent reads on demand, while public libraries now hold thousands of them. Which skill to read has thus become a decision the policy itself makes in the middle of an episode, yet no existing signal trains it. W

Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL

👍 92

Reinforcement learning (RL) has emerged as a powerful approach for improving reasoning in language and vision-language models, yet its strongest successes still depend heavily on ground-truth supervision (e.g., verifiable reward). Such annotations are costly to obtain and become increasingly scarce

Chain-of-Experience for Continual LLM Improvement

👍 8

Humans continuously learn from experience, whereas conventional large language model (LLM) evaluations ignore the models' ability to improve through inference-time interaction. In this paper, we study how LLMs learn from iterative experience at test time, a setting we refer to as Chain-of-Experience

Towards Real-Time and Adaptable LiDAR Scene Completion

👍 3

LiDAR scene completion is a key component of 3D perception in autonomous driving, where the scene must be completed in real time to be usable in downstream tasks. Existing approaches typically follow an initialize-and-refine paradigm, in which a coarse initialization of the scene is first constructe

LLMs Get Smarter from Targeted Synthetic Multilingual Data

👍 4

Language-specific competency (LSC) is the phenomenon of a language model performing better or worse depending on the language of the prompt. In other words, a language model outputs different (and potentially incorrect) responses to the same semantic query when prompted in different languages. Prior

Bounded Agents: Delegation Security for Multi-Agent AI Systems

👍 4

LLM-based agents can act on behalf of a user to access cloud services, call tools, or invoke agents. At session start, the agent's permissions are set but remain static, and each request is evaluated independently, without considering prior actions. Within its permissions, an agent may act contrary

The More Popular, The Harder to Forget: Adaptive Popularity for LLM Unlearning

👍 17

Popular facts are memorised more deeply during pretraining and resist removal longer than rare ones, yet existing LLM unlearning methods apply uniform gradient pressure regardless of training-data frequency. We propose the AdaPop (Adaptive Popularity) method, which combines local token confidence wi

The Embedder's Dilemma: LLMs Are Better, but at What Cost?

👍 11

Should you replace your text-embedding pipeline with a large language model? We answer this with a controlled, cost-aware comparison of ten LLMs across six families and 26 embedding models (118M to 14B parameters) on 37 tasks spanning classification, semantic textual similarity (STS), clustering, pa

QuoteBench: How Matched Scores Can Hide Command-Path Failures

👍 7

LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot distinguish command-generation errors from failures introduced after generation. QuoteBench measures this boundary with exact final-state validation on 5

SkillEvo: Self-Renewing Evolution Gradients from Multi-Turn Interaction Feedback

👍 30

Agent Skills are today either hand-authored or produced in a single LLM generation pass, and consequently possess no closed loop through which they might improve from the interaction failures they actually cause. Recent work does close this loop, but derives its feedback from single-turn question-an

OmniScientist: An Omni-Modal Omni-Discipline AI Scientist

👍 88

Recent advances in foundation models have enabled AI scientists to automate increasingly complete research workflows, from hypothesis generation and code execution to manuscript preparation. Yet workflow coverage alone does not provide access to the full evidence on which scientific discovery depend

Temporal Multi-Signal Fusion for Token-Level Hallucination Detection

👍 3

Token-level hallucination detectors score each token independently from a single signal, and fail exactly when the generating model is confidently wrong. This paper instead treats hallucination as a temporally extended span and detects it by sequence labeling: each token is scored from a 33-dimensio

AI Engineering Skills Map: Building and Deploying AI Applications

@AndrewYNg · 1.8M 粉丝 · 62.4K 阅 · 1.2K 赞 · 189 转

I previously wrote about our AI Engineering Skills Map, with the highest level skills being (i) Building and deploying AI applications, (ii) Software engineering fundamentals, (iii) Using coding

中文介绍 AndrewYNg分享AI工程技能图谱,聚焦于构建和部署AI应用,强调高级技能如软件工程基础和代码使用。

The Evolution of the Agent Harness

Models keep absorbing the harness into their weights — soon, it will be a harness for human attention rather than for the model.

中文介绍 模型吸收人类注意力机制,未来将更关注人类注意力而非模型自身。

Simulation: the new Scaling Law — Joon Sung Park, Simile AI

Simile’s CEO about his journey from the viral Generative Agents to creating 8 Billion Digital Twins of every living human... and why it’s gone from fun exploration to very serious business.

中文介绍 Simile AI CEO分享其从生成代理到创建80亿数字双胞胎的历程,业务从探索转为严肃。

not much happened today

**Ox Alpha** emerged as a mystery model with strong coding and agentic performance, likely a **Zhipu/GLM-family** model such as **GLM-5.3 Vision**. Analysts suggest its gains come from post-training and infrastructure improvements rather than sheer size, based on the **743B base** of **GLM-5.2** wit

中文介绍 Ox Alpha成为神秘模型,表现优异,分析师认为其增长来自训练和基础设施改进。

Debates over AI consciousness are a trap

“Runaway” AI, “rogue” agents, and “autonomous” actors—the current rhetoric would have you believe that AI agents are not only awake and aware, but angry at their creators. Prominent tech leaders such as Demis Hassabis, Dario Amodei, and Sam Altman push for regulation of these seemingly “superhuman”

中文介绍 AI意识辩论陷入误区,技术领袖呼吁监管。

Unlocking hidden revenue streams with market models

Each day, an airline transports tens of thousands of passengers on hundreds of flights. Often these are not straightforward point-to-point routes, with passengers requiring multiple connections. The airline can consider potentially hundreds of variables to price each of these journeys: demand, seaso

中文介绍 市场模型解锁隐藏的收益流。

Introducing AI Futures

Introducing AI Futures, a new OpenAI blog exploring how transformative AI could reshape power, governance, the economy, and individual freedom.

中文介绍 OpenAI推出AI未来博客,探讨AI如何重塑权力、治理、经济和个人自由。

not much happened today

**OpenAI** and **Anthropic** expanded their agent platforms with new desktop features, collaborative editing, and composable APIs like Skills and Files API. **OpenAI** rolled out memory and workflow features in the EEA, UK, and Switzerland. **AT&T** revealed that 40% of employee AI usage routes to o

中文介绍 OpenAI和Anthropic扩展代理平台,推出新桌面功能、协作编辑和可组合API。AT&T透露40%员工...

Australia news live: shark attack victim thanks donors after ‘incredibly difficult’ recovery; second fur seal dies of suspected bird flu

Leah Stewart lost her arm in a great white shark attack at Sydney’s Coogee beach. Follow today’s news live Get our breaking news email, free app or daily news podcast Joyce steps in on One Nation’s migration target Barnaby Joyce defended a fellow One Nation MP over comments suggesting the minor part

中文摘要 悉尼库吉海滩发生大白鲨袭击,Leah Stewart失去手臂。

Burnham to visit Ukraine with promise of boost for Kyiv’s long-range missiles

PM will use independence day trip to announce sharing of classified information about Storm Shadow components Britain will boost Ukraine’s capacity to make its own long-range missiles, Andy Burnham will announce on Monday, as he travels to Kyiv for the country’s independence day on his first foreign

中文摘要 英国首相宣布将向基辅提供风暴阴影导弹技术信息,以增强其制造远程导弹的能力。

Landslide at waste mound in Guinea capital kills 30, government says

Dump site in Conakry, which minister had just promised to move, collapsed after heavy rain in west African state A ⁠landslide at a huge waste dump in Guinea’s capital has killed 30 people, the government said on Sunday, after heavy rains overnight prompted it to ⁠collapse, engulfing nearby tents and

中文摘要 几内亚首都科纳克里一垃圾山发生滑坡,造成30人死亡。

Learning to live in a post-American world

In a recent article for Foregin Affairs, Mark Leonard makes that argument that America's leading role in the world order is fading. He discusses his argument with NPR's Danielle Kurtzleben.

中文摘要 马克·伦纳德在《外交事务》杂志中提出,美国在世界秩序中的领导地位正在衰落。

The Price of Bread

War, climate shocks and trade disputes are pounding the world’s breadbaskets and pushing the global food system into dangerous territory.

中文摘要 战争、气候冲击和贸易争端正在打击全球粮食产区,将全球食品体系推向危险境地。

Is Israel about to split the occupied West Bank in half?

Israel is moving forward with its ‘E1’ plan. What is it and why could it threaten the future of a Palestinian state?

中文摘要 以色列推进“E1”计划,可能威胁巴勒斯坦国的未来。

Trump waves green flag to start IndyCar race through Washington streets

The president took a ceremonial lap of the track before the race, which is the culmination of a summer of events to mark the country's 250th birthday.

中文摘要 特朗普为华盛顿街道上的IndyCar比赛挥舞绿色旗帜,庆祝国家250岁生日。

Severe winds toss four aircraft across Italian airport tarmac

Winds reaching around 120 km/h swept through Forlì in Italy’s Emilia-Romagna region, overturning four light aircraft.

中文摘要 意大利艾米利亚-罗马涅地区Forlì的强风将四架轻型飞机吹翻。

Why students are being paid £2,000 to play computer games

Roehampton University offers students £2,000 a year to play esports alongside their studies.

中文摘要 罗汉普顿大学为学生在学习的同时玩电子竞技游戏提供每年2000英镑的报酬。

Scottish state bank records £138mn loss on back of failed investments

Fifth consecutive full-year loss exposes tension in mandate of lender set up by SNP-led government

中文摘要 苏格兰国家银行连续第五年录得1380万英镑亏损,暴露了由SNP领导政府设立的贷款机构职能上的紧张关系。

Burnham Enters World Stage With Debut Foreign Trip to Ukraine

Andy Burnham is traveling to Ukraine on Monday for his first international trip as British prime minister, in an attempt to show he’ll maintain diplomatic and military support for Kyiv as Russia’s war enters a decisive phase.

中文摘要 英国首相Burnham首次访问乌克兰,试图展示英国将继续对基辅提供外交和军事支持。

Hong Kong Homebuyers Left Cold by Prices in Far-Flung Tech Hub

Hong Kong’s ambitious plan to create a technology hub near the mainland Chinese border faces an early test after the first residential sales in its Northern Metropolis got off to a middling start, highlighting the hurdles to turning the area into a new economic engine.

中文摘要 香港在边境附近建立科技园区的雄心计划面临初步挑战,北部都会区首次住宅销售表现平平,凸显将地区转变为新经济引擎的障碍。

Taiwan Dollar’s August Rebound Is Looking Fragile, Analysts Say

The Taiwan dollar’s strongest monthly rise in over a year is running into a wall of headwinds, sowing doubts among analysts on the sustainability of its rebound.

中文摘要 分析师表示,台币在8月份强劲反弹后,正面临重重困难,引发对其反弹可持续性的怀疑。

Hedge Fund Manager Phil King to Retire From Regal Partners

Regal Partners said Phil King will start a transition to retirement, marking the exit of one of Australia’s most high-profile hedge fund managers.

中文摘要 Regal Partners表示,基金经理Phil King将开始退休过渡,标志着澳大利亚最著名对冲基金经理之一的离职。

Shein Seeks Up to $1.8 Billion in Long-Awaited Hong Kong IPO

Shein Global Holdings Ltd. is seeking to raise as much as HK$13.9 billion ($1.8 billion) in its Hong Kong initial public offering, as it enters the final stretch of an arduous journey to go public.

中文摘要 Shein Global Holdings Ltd. 正在寻求通过其香港首次公开募股筹集高达139亿港元(18亿美元),以完成其艰难的上市之旅。

Oil Dips, Alibaba Funding Plan Keeps AI in Focus: Markets Wrap

Oil edged lower in at the start of a key week for markets, with traders focused on the US plan to economically isolate Iran and the Federal Reserve’s annual gathering.

中文摘要 油价小幅下跌,阿里巴巴的资金计划让AI成为市场关注焦点。

FirstFT: Kevin Warsh seeks to calm investors’ nerves as signs of economic strain grow

Also in this newsletter: Alibaba announces $10.2bn share placement and Anthropic’s best AI model struggles to attract users

中文摘要 凯文·沃什寻求安抚投资者情绪,同时经济压力迹象日益增多。此外,阿里巴巴宣布发行102亿美元股份,Anthropic的顶级AI模型难以吸引用户。

How New Zealand’s biggest city solved its housing crisis

Rezoning has changed the face of Auckland though critics warn lessons hard to replicate

中文摘要 新西兰最大城市通过重新规划改变了面貌,但批评者警告说,这些经验难以复制。

该源今日无内容。