每日简报

2026-09-03

← 历史归档

fmtlib/fmt

C++ · ★ 24,226 · 🍴 2,984 · 📈 14 stars today

A modern formatting library

中文介绍 fmt 是一个现代格式化库,用于字符串和数字的格式化,适用于 C++ 项目。

google-research/timesfm

Python · ★ 29,744 · 🍴 2,860 · 📈 343 stars today

TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting.

中文介绍 TimesFM 是由谷歌研究团队开发的预训练时间序列基础模型,用于时间序列预测。

DietrichGebert/ponytail

JavaScript · ★ 121,476 · 🍴 6,567 · 📈 1,354 stars today

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

中文介绍 ponytail 是一个使你的 AI 代理表现得像最懒散的老司机的工具,通过避免编写不必要的代码来优化代码质量。

debpalash/VoiceStudio

Python · ★ 14,668 · 🍴 2,101 · 📈 832 stars today

VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

中文介绍 VoiceStudio 是一个开源的语音克隆、设计、视频配音、语音转写、转录和有声书制作工具,支持646种语言。

sngyai/Sequoia-X

Python · ★ 6,030 · 🍴 1,244 · 📈 63 stars today

A股自动选股系统 — 多种技术形态自动扫描,收盘后自动运行并推送飞书

中文介绍 Sequoia-X 是一个A股自动选股系统,使用多种技术形态自动扫描,并在收盘后自动运行并推送飞书消息。

ChromeDevTools/chrome-devtools-mcp

TypeScript · ★ 50,626 · 🍴 3,553 · 📈 148 stars today

Chrome DevTools for coding agents

中文介绍 Chrome DevTools for coding agents 是为编码代理设计的 Chrome 开发者工具。

NousResearch/hermes-agent

Python · ★ 240,111 · 🍴 49,122 · 📈 533 stars today

The agent that grows with you

中文介绍 Hermes-Agent 是一个与用户成长同步的代理,旨在提高人工智能的交互体验。

superlinked/sie

Python · ★ 3,048 · 🍴 299 · 📈 60 stars today

Open-source inference server and production cluster for all the models your agent needs.

中文介绍 Sie 是一个开源的推理服务器和生产集群,适用于所有你的代理需要的模型。

pacifio/atlas

Rust · ★ 2,857 · 🍴 186 · 📈 888 stars today

Source control for agents. Use multiple coding agents, track their changes and query them in one place

中文介绍 Atlas 是一个为代理提供的源代码控制工具,可跟踪多个编码代理的更改并在一个地方查询它们。

zyronon/TypeWords

Vue · ★ 9,292 · 🍴 1,129 · 📈 21 stars today

Practice English, one strike, one step forward; 练习英语,一次敲击,一点进步;

中文介绍 TypeWords 是一个英语练习工具,通过逐步增加难度来帮助用户提高英语水平。

Imbad0202/academic-research-skills

Python · ★ 45,544 · 🍴 3,580 · 📈 799 stars today

Academic Research Skills for Claude Code: research → write → review → revise → finalize

中文介绍 Academic Research Skills for Claude Code 提供了从研究到最终定稿的学术研究技能指南。

affaan-m/ECC

JavaScript · ★ 246,317 · 🍴 37,146 · 📈 516 stars today

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

中文介绍 ECC 是一个性能优化系统,旨在为 Claude Code、Codex、Opencode、Cursor 等提供技能、本能、记忆、安全和研究优先的开发。

protocolbuffers/protobuf

C++ · ★ 71,933 · 🍴 16,271 · 📈 18 stars today

Protocol Buffers - Google's data interchange format

中文介绍 Protocol Buffers 是 Google 的数据交换格式,用于序列化结构化数据。

vercel-labs/portless

TypeScript · ★ 11,738 · 🍴 383 · 📈 73 stars today

Replace port numbers with stable, named local URLs. For humans and agents.

中文介绍 Portless 是一个替代端口号的解决方案,使用稳定的本地 URL 替代端口号,适用于人类和代理。

blader/humanizer

Python · ★ 40,348 · 🍴 3,493 · 📈 374 stars today

Agent skill that removes signs of AI-generated writing from text

中文介绍 Humanizer 是一个 AI 写作技能,用于从文本中移除 AI 生成的迹象。

JuliusBrussee/caveman

Go · ★ 102,623 · 🍴 5,970 · 📈 238 stars today

🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman

中文介绍 Caveman 是一个 Claude Code 技能,通过使用类似穴居人的语言来减少65%的标记数量。

mattpocock/skills

Shell · ★ 245,206 · 🍴 20,849 · 📈 1,166 stars today

Skills for Real Engineers. Straight from my .agents directory.

中文介绍 Skills for Real Engineers 提供了真实工程师所需技能,直接从 .agents 目录中提取。

Gitlawb/openclaude

TypeScript · ★ 31,968 · 🍴 8,994 · 📈 775 stars today

runs anywhere. uses anything

中文介绍 OpenClaude 是一个运行在任何地方、使用任何工具的开源项目。

firecrawl/pdf-inspector

Rust · ★ 18,491 · 🍴 1,245 · 📈 586 stars today

Fast Rust library for PDF inspection, classification, and text extraction. Intelligently detects scanned vs text-based PDFs to enable smart routing decisions.

中文介绍 PDF Inspector 是一个快速 Rust 库,用于 PDF 检查、分类和文本提取,智能地检测扫描和基于文本的 PDF,以实现智能路由决策。

Knowledge Distillation During Mid-Training Favors Reasoning over Factual Recall

👍 1

Logit-based knowledge distillation (KD) is used to train smaller language models (LMs) via supervision from stronger teachers, but whether its benefits are consistent across training stages remains unclear. Through controlled experiments, we find that forward Kullback-Leibler (KL) distillation--the

DramaChain Bench: An End-to-End Benchmark for Short-Drama Generation

👍 2

Commercial short-drama production follows a multi-stage chain: script, storyboard, keyframe imagery, shot-level video, and the finished short drama. Most existing benchmarks evaluate solely the video-generation stage using pre-authored inputs instead of real upstream pipeline outputs. This leaves tw

DiagEvo: Diagnosis-Guided Self-Evolution via Hierarchical Error Memory

👍 14

Self-play is an effective paradigm for language-model self-evolution, but without guidance, solver performance can plateau or decline across rounds. Unguided methods steer question generation with signals such as difficulty, learnability, or diversity. These signals keep questions challenging and va

StudentSim: Training LLM-based Student Simulators

👍 458

AI tutors are most useful when they adapt to each student's strengths, weaknesses, and preferred guidance, but evidence about which guidance works for which student is sparse, slow, and costly to collect from real learners. Student simulators can provide this signal as a proxy, yet existing approach

Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs

👍 7

Prompt optimization can improve multi-agent LLM systems, but the prompts being optimized often serve two entangled roles: generating task-relevant content and specifying execution-critical protocols, such as message routing, output formatting, and termination signals, on which the underlying code re

EM^2Mem: Event-Centric Multimodal Memory for Large Language Models

👍 9

Multimodal memory offers a scalable interface for long-video question answering, but existing methods often retrieve captions, frames, transcripts, summaries, or graph facts as isolated fragments. Although searchable, such fragments are not generation-ready: language models must reconstruct cross-mo

The Safeguard Worked. Is the LLM System Safer?

👍 3

Safeguards in deployed LLM services are evaluated by refusal, attack success, and policy violation rates. Those rates characterize how a control performed on the requests it was tested on. A deployment has to answer a different question: how much help with harmful tasks the service still gives an at

Agents in the Large: Perception-Centered Architecture for Persistent Agents

👍 8

Cognitive language agents have achieved substantial progress by equipping language models with memory, tools, and decision-making procedures, enabling agents to reason and act in interactive environments. Existing frameworks largely cast these agents as systems for solving user-specified, bounded ta

Verification-Aware Training for Speculative Decoding

👍 5

Speculative decoding accelerates large language model inference by using a draft model to generate candidate tokens, which are verified by the target model in a single forward pass. Verification proceeds sequentially and discards every position from the first rejection onward, yet existing draft tra

CoVA-SFT: A Large-Scale Dataset for Chain of Visual Abstractions

👍 4

Chain-of-thought (CoT) reasoning has dramatically improved large language models (LLMs) by allowing them to decompose problems into intermediate steps. While CoT is widely effective for linguistic tasks, text-only CoT forces models to serialize visual problems into awkward prose. Although architectu

EvoGenUI-Bench: Evaluating LLMs as Multi-Turn Generative UI Assistants

👍 4

Large language models can generate interactive web interfaces, but reliable generative UI requires maintaining an executable artifact as user requests evolve. We introduce EvoGenUI-Bench, a benchmark for multi-turn interface maintenance comprising 150 five-turn tasks and 750 turns across three scena

Chat-Edit-3D++: Interactive 3D and 4D Scene Editing via Large Language Models

👍 3

Recent work on image content manipulation based on vision-language pre-training models has been effectively extended to text-driven 3D scene editing. However, existing schemes for 3D scene editing still have certain shortcomings, hindering their further development as interactive design tools. Such

Evaluating the Hidden Costs of Personalization in Large Language Models

👍 25

While Large language models (LLMs) incorporate user personalization signals to improve usability and helpfulness, they increasingly shift from providing balanced, informative responses toward optimizing for user satisfaction when conditioned on personal context such as conversation history, inferred

UI-Venus-2 Technical Report

👍 55

Multimodal GUI agents have emerged as a promising paradigm for digital task automation, yet transitioning from benchmark-oriented models to dependable real-world applications remains challenging due to limited environment coverage, brittle task construction, and unreliable reward verification. In th

Nobody is talking seriously about AI demand

@giovannicatt3 · 1.4K 粉丝 · 142.0K 阅 · 507 赞 · 56 转

AI capex for 2028 is forecast to be larger than the budget of France. Frontier AI labs’ revenue ramp justifies almost any number. Reflexivity in AI demand is a double-edged sword, and it's now a good

中文介绍 分析AI需求预测,指出2028年AI资本支出可能超过法国预算,强调AI需求自反性是一把双刃剑。

Compilers 2.0: AI as stochastic optimizer

@cdleary · 2.3K 粉丝 · 100.3K 阅 · 533 赞 · 66 转

There has been a lot of discussion following the presentation of the Jalapeño MLA kernel at HotChips and subsequent commentary by SemiAnalysis. As OpenAI’s hardware team, we just barely touched on

中文介绍 探讨Jalapeño MLA内核在HotChips上的展示及其后续评论,提及OpenAI硬件团队对AI随机优化器的初步探讨。

Department of War Launches OpenAI's ChatGPT Mil on GenAI.mil

@DoWCTO · 104.9K 粉丝 · 86.1K 阅 · 663 赞 · 139 转

The Department of War today launched OpenAI's ChatGPT Mil, marking it as the next frontier AI capability housed on GenAI.mil, the Department's premier generative AI platform. This

中文介绍 美国国防部推出OpenAI的ChatGPT Mil,作为GenAI.mil平台上的下一代AI能力。

Agentic Engineering Setup (after 2,000+ hours)

@DavidOndrej1 · 67.4K 粉丝 · 68.7K 阅 · 513 赞 · 40 转

Over the past 3 years, I've spent well over 2,000 hours coding with AI, and I've personally interviewed some of the most productive people in the space of Agentic Engineering. Below is my full Agentic

中文介绍 分享超过2000小时AI编码经验,介绍Agentic Engineering的设置和高效实践。

Grok Bot for Engineering

@lingxi · 7.0K 粉丝 · 64.4K 阅 · 629 赞 · 44 转

I’m a SpaceXAI engineer building Grok Bot with Grok Bot. Think of Grok Bot as a highly capable engineering intern, with its own computers, that can manage coding agents and learn from how you work. It

中文介绍 介绍Grok Bot,一款具备自身计算机,可管理编码代理并学习用户工作方式的工程助手。

7 Grok Bots for Marketing

@irabukht · 18.4K 粉丝 · 63.3K 阅 · 536 赞 · 28 转

18 things to know and 7 paste-in prompts for running ads, SEO and GEO on Grok Bot. Collected from 1,000+ marketers in our community. Most of us use Grok Bot to run marketing work end to end: Weekly

中文介绍 提供7个Grok Bot营销使用提示,收集自1000多位营销人员,涵盖广告、SEO和GEO等全流程工作。

Department of War Launches Starshield AI's Grok for Government on GenAI.mil

@DoWCTO · 104.9K 粉丝 · 57.5K 阅 · 924 赞 · 174 转

The Department of War today launched Starshield AI's Grok for Government, expanding the suite of frontier AI capabilities on GenAI.mil, the Department's premier generative AI platform.

中文介绍 美国国防部在GenAI.mil平台上推出Starshield AI的Grok for Government,扩展前沿AI能力。

Your AGENTS.md is holding you back

@posthog · 24.3K 粉丝 · 43.7K 阅 · 503 赞 · 33 转

Context engineering used to focus on adding information that base models lacked: But then the models kept getting better. Now, the same context your agents couldn’t function without can make them

中文介绍 指出基础模型不断进步导致上下文工程的重要性降低,强调上下文可能成为限制因素。

Introducing agentic video understanding with Gemini

@GoogleAIStudio · 197.2K 粉丝 · 37.4K 阅 · 522 赞 · 48 转

Our new agentic feature for video analysis cuts costs by up to 66% and reduces token consumption by up to 88% while boosting accuracy. Today, we’re launching agentic video understanding across our

中文介绍 推出Gemini视频分析的新功能,降低成本并提高准确性,同时减少token消耗。

Using Grok Bot: 8 templates to get inspired

@mattyp · 57.4K 粉丝 · 29.7K 阅 · 508 赞 · 40 转

Templates are now available for Grok Bot, so I wanted to share my favorites. Templates contain the skills, memories, and official @bot plugins from your bots. When you use a template, you get a copy

中文介绍 分享Grok Bot的8个模板,包含技能、记忆和官方插件,以激发灵感。

A Lagrangian View of Flow Matching

@docmilanfar · 115.8K 粉丝 · 26.0K 阅 · 512 赞 · 56 转

This is a pedagogical post meant to offer an intuitive, bottom-up perspective on Flow Matching. If you have ever wondered why some models require hundreds of iterative steps to generate an image while

中文介绍 从拉格朗日视角介绍Flow Matching,解释为何某些模型需要数百次迭代步骤生成图像。

From rg to zg: Local Search Beyond Keywords

@QwenDevs · 10.9K 粉丝 · 21.4K 阅 · 509 赞 · 55 转

Summary: The information humans and agents need is often scattered across large numbers of local files, making it difficult to locate accurately and efficiently. zg (zvec-grep) is local-first search

中文介绍 介绍zg(zvec-grep)作为本地搜索工具,解决人类和代理在大量本地文件中查找信息的问题。

Facilitating AI integration with simplicity at scale

As companies scale, the technology supporting operations can become a liability just as quickly as it becomes an asset. Disconnected systems, site-specific tools, spreadsheets, and manual workarounds can create data silos that make it harder to spot problems early, coordinate responses, and make dec

中文介绍 文章讨论公司规模扩大时,如何简化AI集成,避免技术成为负担。

ATV Big Air Tour turned 3 days of work into 3 hours with ChatGPT

ATV Big Air Tour uses ChatGPT Work to speed up marketing, merchandising, and more. It even turned merchandise photos into an inventory website in 15 minutes.

中文介绍 OpenAI报道ChatGPT如何将ATV Big Air Tour的工作效率提高至原来的1/3。

How AI-native companies turn workflows into operating capability

Basis, Clay, and Exa Labs use AI agents to improve onboarding, account management, and developer integrations. See what enterprise leaders can apply.

中文介绍 OpenAI介绍AI原生公司如何将工作流程转化为运营能力。

Path to Astra: critical capabilities and frontier safeguards

Astra is the first OpenAI model to meet the Critical cybersecurity capability threshold under the Preparedness Framework, with stronger safeguards for release.

中文介绍 OpenAI的Astra模型满足关键网络安全能力阈值,具备更强大的安全防护。

Claude Fable 5.1 and Claude Mythos 5.1

**Anthropic** released **Claude Fable 5.1** and **Claude Mythos 5.1**, which share base weights but differ in safeguards and routing, showing improved coding performance and usability with a **75% cache-read price cut to $0.25/MTok**. Benchmarks highlight strong coding/science results, though Fable

中文介绍 Anthropic发布Claude Fable 5.1和Claude Mythos 5.1,基准测试显示编码和科学能力增强。

In Small Iranian Town, a U.S. Attack Turns a Wedding Into a Tragedy

The bomb that hit the residential area was American, according to a weapons expert and a visual analysis by The Times. It killed five people and wounded 67 others, the Iranian authorities said.

中文摘要 美国空袭伊朗一住宅区,造成5人死亡,67人受伤。

Australia news live: Marles and Hegseth meet in Washington; teen charged after ‘near complete amputation’ of boy’s hand

Boy to face court today. Follow the day’s news live. Get our breaking news email, free app or daily news podcast Shadow treasurer says tobacco plans would bring in extra billions and stop people buying from the ‘black market’ Tim Wilson, the shadow treasurer, said the tobacco proposal would bring in

中文摘要 澳大利亚副财长表示烟草计划将带来数十亿额外收入并遏制黑市交易。

Dutch central bank moves 86 tonnes of gold to UK from US and Canada, citing ‘geopolitical unrest’

Bank says gold reserves held in London could be traded more easily and the move will allow it to respond more rapidly in a ‘crisis situation’ The Dutch central bank says it has moved 86 tonnes of its gold reserves out of the US and Canada to London, citing “increasing geopolitical unrest”. De Nederl

中文摘要 荷兰央行将86吨黄金从美国和加拿大转移到英国,以应对地缘政治动荡。

Did DOGE ruin America’s food safety system?

Food safety scares raise questions over whether Trump’s cuts weakened US food safety monitoring.

中文摘要 特朗普削减食品安全监管引发对美国食品安全系统的担忧。

China’s falling emissions amid Iran war spark hope of decarbonisation watershed

Oil consumption plummets and EV sales soar as analysts say demand may not fully return even if crude price falls China’s carbon dioxide emissions fell by 1% after the outbreak of the US-Israeli war on Iran, thanks to a sharp reduction in oil consumption and a steady rise in the use of electric vehic

中文摘要 伊朗战争爆发后,中国二氧化碳排放量下降1%,分析师表示石油消费减少,电动汽车销量增加。

Andy Burnham targets migrant crisis ahead of Macron’s visit to No 10

PM will aim to centre talks with French president on key foreign policy issues such as Iran, smuggling gangs and EU relations Andy Burnham will begin the process of courting European leaders when he welcomes Emmanuel Macron as the first foreign head of state to visit him in Downing Street. The prime

中文摘要 英国首相 Burnham 将与法国总统 Macron 谈论伊朗、走私团伙和欧盟关系。

Apple Maps renames Lake Ontario as ‘Lake America’ for US users after Trump order

Tech company submits to controversial Trump order to rename body of water amid US trade spat with Canada Apple has renamed Lake Ontario as “Lake America” for US users of its Maps app, after an executive order from Donald Trump to change the name of the Great Lake amid his trade spat with Canada. The

中文摘要 苹果公司将安大略湖更名为“美洲湖”,以符合特朗普政府的命令。

Maduro Asks US Court to Dismiss Charges in Drug-Trafficking Case

Nicolás Maduro asked a judge in New York to throw out a criminal case against him, arguing that he’s immune to prosecution in the US as the head of a sovereign country.

中文摘要 尼科拉斯·马杜罗要求纽约法院驳回对其毒品走私案件的指控,称作为主权国家领导人,他在美国享有豁免权。

Japan’s 30-Year Bond Auction Risks Adding Fuel to Debt Selloff

Japan’s 30-year government bond auction Thursday will test investor appetite as a global selloff pushes long-dated yields to their highest levels in almost two decades.

中文摘要 日本周四将进行的30年期国债拍卖可能加剧债务抛售,全球抛售推动长期收益率达到近20年来最高水平。

China’s Waning Appetite for Oil Is Keeping Emissions in Check

The millions of Chinese drivers who turned to charging stations over gas pumps as the Iran War drove up fuel prices have helped shift the country’s climate trajectory.

中文摘要 由于伊朗战争推高燃料价格,数百万中国司机转向充电站而非加油站,这有助于改变国家的气候轨迹。

Five Below Lifts Outlook, Aims to Ramp Up Appeal to Gen Alpha

Five Below Inc. shares rose over 8% in postmarket trading on Wednesday after the retailer lifted its full-year outlook, citing the addition of new stores and the popularity of its budget-friendly merchandise among Gen Alpha and Gen Z shoppers.

中文摘要 Five Below Inc.股价在周三盘后交易中上涨超过8%,该公司上调了全年展望,称新增门店和预算友好型商品在千禧一代和Z世代购物者中的受欢迎程度增加。

Steve Ballmer banned by NBA over improper payment to basketball star

LA Clippers team funnelled millions of dollars to star Kawhi Leonard to circumvent league salary cap, league investigation finds

中文摘要 洛杉矶快船队因向球星科怀·伦纳德支付数百万美元以规避联盟薪资上限而被NBA禁止,联盟调查发现。

该源今日无内容。