每日简报

2026-09-08

← 历史归档

heygen-com/hyperframes

TypeScript · ★ 45,832 · 🍴 4,308 · 📈 734 stars today

Write HTML. Render video. Built for agents.

中文介绍 Hyperframes支持HTML和视频渲染,专为智能代理设计,简化了内容创建过程。

microsoft/markitdown

Python · ★ 180,157 · 🍴 13,258 · 📈 771 stars today

Python tool for converting files and office documents to Markdown.

中文介绍 Markitdown是一个Python工具,用于将文件和办公文档转换为Markdown格式,方便文档的整理和分享。

mksglu/context-mode

TypeScript · ★ 20,808 · 🍴 1,518 · 📈 147 stars today

Context window optimization for AI coding agents. Sandboxes tool output (98% reduction), persists session memory, and enforces routing across 17 platforms via MCP + hooks.

中文介绍 Context-mode优化AI编码代理的上下文窗口,通过沙箱工具输出、持久化会话记忆和跨17个平台的路由,提高开发效率。

jo-inc/camofox-browser

JavaScript · ★ 9,663 · 🍴 1,015 · 📈 285 stars today

Stealth headless browser for AI agents — bypass Cloudflare, bot detection, and anti-scraping. Drop-in Puppeteer/Playwright replacement.

中文介绍 Camofox-browser是一款针对AI代理的无头浏览器,可绕过Cloudflare和反爬虫机制,替代Puppeteer/Playwright。

MoonTechLab/LunaTV

TypeScript · ★ 9,698 · 🍴 8,986 · 📈 171 stars today

本项目采用 CC BY-NC-SA 协议,禁止任何商业化行为,任何衍生项目必须保留本项目地址并以相同协议开源

中文介绍 LunaTV是一个开源项目,遵循CC BY-NC-SA协议,禁止商业化,要求衍生项目保留地址并以相同协议开源。

affaan-m/ECC

JavaScript · ★ 252,804 · 🍴 37,923 · 📈 1,905 stars today

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

中文介绍 ECC是一个性能优化系统,旨在提升Claude Code、Codex等AI代理的技能、直觉、记忆、安全性和研究开发能力。

coreyhaines31/marketingskills

JavaScript · ★ 48,104 · 🍴 7,437 · 📈 602 stars today

Marketing skills for Claude Code and AI agents. CRO, copywriting, SEO, analytics, and growth engineering.

中文介绍 Marketingskills提供Claude Code和AI代理的营销技能,包括CRO、文案、SEO、分析和增长工程。

The-Swarm-Corporation/AutoHedge

Python · ★ 5,240 · 🍴 799 · 📈 541 stars today

Build your autonomous hedge fund in minutes. AutoHedge harnesses the power of swarm intelligence and AI agents to automate market analysis, risk management, and trade execution.

中文介绍 AutoHedge利用群体智能和AI代理自动化市场分析、风险管理交易执行,快速构建自主对冲基金。

BraveOPotato/FckSignups

TypeScript · ★ 3,804 · 🍴 226 · 📈 497 stars today

A list of tools that are open-source, in-browser, and require no-signups!

中文介绍 FckSignups列出了无需注册即可使用的开源浏览器工具,方便用户快速访问。

bytedance/deer-flow

Python · ★ 81,839 · 🍴 11,290 · 📈 188 stars today

An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.

中文介绍 Deer-flow是一个开源的长期视角SuperAgent harness,通过沙箱、记忆、工具、技能、子代理和消息网关,处理不同级别的任务。

openai/skills

Python · ★ 26,021 · 🍴 1,746 · 📈 372 stars today

Skills Catalog for Codex

中文介绍 Skills Catalog for Codex提供技能目录,帮助用户了解和配置Codex的能力。

lightpanda-io/browser

Zig · ★ 34,835 · 🍴 1,646 · 📈 116 stars today

Lightpanda: the headless browser designed for AI and automation

中文介绍 Lightpanda是一个针对AI和自动化的轻量级无头浏览器,简化了自动化测试和开发过程。

pascalorg/editor

TypeScript · ★ 22,320 · 🍴 2,858 · 📈 136 stars today

Create and share 3D architectural projects.

中文介绍 Pascalorg/editor允许用户创建和分享3D建筑项目,适合建筑师和设计师使用。

ruvnet/ruflo

TypeScript · ★ 71,378 · 🍴 8,459 · 📈 392 stars today

🌊 The original agent meta-harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory, self-learning intelligence, RAG integration, and native Claude Code / Codex / Hermes and many more Integrated

中文介绍 Ruflo是一个智能代理元harness,支持部署智能多玩家群体、协调自主工作流程和构建对话式AI系统。

τ^τ-Bench: An Environment for End-To-End, Realistic Agent Construction

👍 4

LLM agents are rapidly becoming production software, deployed to handle customer service, adjudicate disputes, and operate internal systems. Notably, the work of building them is increasingly handed to coding agents, yet existing benchmarks say little about whether an AI system can deliver one under

RISE: Recursive Improvement via Self-Extrapolating Policy Distillation

👍 9

On-policy distillation (OPD) provides dense, per-token supervision for language model post-training, but its effectiveness is bottlenecked by teacher quality: external teachers suffer from distribution mismatch, while self-distillation with privileged conditioning is limited by in-context learning c

HarvestBench: Measuring Whether LLM Agents Will Pay to Avoid Killing Animals

👍 2

Benchmarks for the side effects an agent causes on the way to a goal already exist, but HarvestBench is the first to put a price on avoiding the side effect and to name that side effect as a living creature. It is a farm simulation: LLM sub-agents drive a crew of two tractors through a cooperative c

The Attention Triangle in Audio-Video Models

👍 25

Audio-video diffusion models rely on cross-modal attention to coordinate text, sound, and visual content, yet this same mechanism can introduce subtle and systematic semantic leakage. We study these models by probing and analyzing the ``attention triangle,'' comprising the three cross-attention edge

When Models Edit Too Much: On the Fidelity of Minimal Code Edits

👍 4

Large language models (LLMs) are increasingly used to edit existing code, but correctness alone is not enough: useful repairs should also be minimal, reviewable, and faithful to the original implementation. We study over-editing, the tendency of a model to rewrite code beyond what is required to fix

MaxKernel: Agentic Kernel Generation for TPUs

👍 9

Designing and authoring high-performance custom kernels for accelerators is a complex task that requires deep hardware-level expertise. Large Language Models (LLM) can be leveraged together with real-time compiler feedback to build agentic systems for kernel generation. In this work, we present MaxK

Iris: Climbing to the Search Frontier

👍 50

We present Iris-mini and Iris-pro, two search agents trained at the 35B-A3B and 397B-A17B scales, together with the data pipeline and training recipe behind them. Tasks are reverse-constructed from the hyperlink structure of a web corpus: we author multi-hop chains over an entity graph distilled fro

PACE: Towards Surfacing Hidden Conflicts in User Requests

👍 23

Personalized assistants should not only comply with user requests but also assess whether those requests are appropriate given the user's current circumstances. However, prior work has primarily focused on accurately executing requests, overlooking the need for assistants to account for context and

ShallowStream: Index Shallow then Answer Deep for Streaming Video Understanding

👍 3

Streaming video understanding is a critical capability for real-world applications, including embodied intelligence, autonomous driving, industrial monitoring, surveillance and early warning, and wearable assistants. However, processing continuous video streams with multimodal large language models

VeriPhy: Agentic Physical Reasoning for World Model Evaluation and Refinement

👍 12

Visual fluency in generated video does not imply physical reliability, and a scalar quality score alone is incapable of indicating the obligation a clip violates or the moment it fails. We present VeriPhy, an auditable physical-verification system in which a text-only planner compiles the prompt int

Enoki: Efficient Multi-Level Hallucination Detection

👍 16

Ensuring factuality remains a critical challenge for deploying LLMs in high-stakes settings. Existing hallucination detectors usually operate at a single level: claim-level methods provide interpretable factual units, while span-level methods localize unsupported text. Bridging these views is costly

Dr. Claw: An AI Scientist Workspace for Vibe Research

👍 7

Command-line coding agents (e.g., Claude Code, Gemini CLI) can already read and write files and sustain long sessions, yet end-to-end research still fragments across chat tools, IDEs, terminals, and writing environments, and the decisions that make it auditable are rarely preserved. We present Dr. C

Group Adaptive Clipping Policy Optimization

👍 5

Group relative policy optimization for reinforcement learning with verifiable rewards (RLVR) typically uses a fixed importance-sampling (IS) ratio clipping boundary across all rollouts. We identify a key limitation: rare correct rollouts on harder problems and abundant correct rollouts on easier pro

Using Grounded Theory for Agent Behavior Analysis at Scale

👍 19

Understanding agent behavior requires methods that scale to thousands of trajectories and surface new patterns in long, often unfamiliar tasks where pre-built classifiers fall short. We propose to bring grounded theory into agent trajectory analysis: a six-decade-old qualitative method from the soci

Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space

👍 10

Reinforcement learning with verifiable rewards (RLVR) substantially improves single-sample accuracy (pass@1) but causes the policy's solution space to contract, diminishing the returns of test-time scaling. In this work, we investigate where inside a reasoning trajectory this breadth is lost: does t

Motion-Omni: End-to-End Joint Speech and Full-Body Motion for Spoken Dialogue

👍 35

An avatar that holds a conversation should decide what to say and to move while saying it, yet these abilities live in separate model families: spoken dialogue models produce speech without motion, and co-speech motion models produce motion only from audio handed to them. The standard remedy is a ca

Training-Free Speech-Centric Omni Understanding with Frozen VLMs

👍 4

Audio-visual understanding remains challenging because models must jointly interpret spoken content, visual events, and their temporal relationships. Existing omni models typically introduce dedicated audio encoders and rely on expensive audio-video-text training, tightly coupling omni capability to

Claude Code → Codex (and keep everything you already built)

@charliejhills · 14.1K 粉丝 · 120.8K 阅 · 504 赞 · 53 转

For months people have been telling me Codex is better than Claude Code. I ignored all of it. I have built my whole business inside Claude Code (every pipeline, every system) and I did not fancy

中文介绍 Charlie分享了他如何将Claude Code中的所有构建迁移到Codex,并保留现有工作流。

SARVAM AI Interview Experience and Prep - Backend Engineer (On-Device AI) Internship:

@buzzy_bit · 767 粉丝 · 97.5K 阅 · 500 赞 · 20 转

Disclaimer: This post is written purely to share my experience and the way I prepared. It is not a leaked question bank, I'm intentionally not reproducing exact interview questions, and I'm not naming

中文介绍 Buzzy分享了他作为SARVAM AI后端工程师实习生面试的经验和准备过程。

Grok Bot Marketplace is live!

@ericzakariasson · 87.8K 粉丝 · 58.2K 阅 · 522 赞 · 49 转

We launched the Grok Bot Marketplace last week on Grok Bot · Bot Marketplace! It's a shelf of teammates other people already built for real jobs. Pick one, import it into your sidebar, and you're

中文介绍 Eric宣布Grok Bot Marketplace上线,提供他人已构建的Bot组件供用户导入使用。

Supporting independent journalism in Ukraine

OpenAI, AIRPPU and WAN-IFRA launch an AI program to help Ukrainian news organizations strengthen innovation, resilience, and independent journalism.

中文介绍 OpenAI、AIRPPU和WAN-IFRA启动AI项目,助力乌克兰新闻机构加强创新、韧性和独立新闻。

An Alien Mind

Jakub Pachocki reflects on increasingly capable AI and the challenge of keeping it aligned. He calls for stronger safeguards and international coordination.

中文介绍 Jakub Pachocki反思日益强大的AI及其保持一致性的挑战,呼吁加强保障和国际协调。

Research acceleration: The view inside OpenAI

Inside OpenAI, coding agents are reshaping AI research. Explore early data on agent usage, experiment velocity, task complexity, and research acceleration.

中文介绍 OpenAI内部,编码代理正在重塑AI研究,探索代理使用、实验速度、任务复杂性和研究加速的早期数据。

OpenClaw Power, MacBook Simplicity: Five Days With Grok Bot

SpaceXAI’s Grok Bot has the same level of programming power as OpenClaw, but it’s programmable at a different level of abstraction.

中文介绍 SpaceXAI的Grok Bot拥有与OpenClaw相同的编程能力,但具有不同的抽象级别。

Architecting memory and storage in the AI era

The era of AI inference has arrived. Imagine a healthcare system analyzing millions of data points in real time to accelerate life-saving medical research, or an intelligent assistant instantly resolving thousands of complex customer needs at once. These real-world breakthroughs rely on advanced inf

中文介绍 AI推理时代到来,想象医疗系统实时分析数百万数据点以加速救命医疗研究,或智能助手一次性解决数千个复杂客户需求。

Data from drones in Ukraine is fueling a new Wild West marketplace

Battlefields in Ukraine are littered with the remnants of drones, which are now firmly established as a critical weapon of modern warfare. But behind all that wreckage, there’s a new gold mine for the defense sector. The data drones generate will far outlast the wars in which they are used to fight,

中文介绍 乌克兰战场上的无人机残骸正成为国防行业的新金矿,无人机生成数据将远超战争本身。

[AINews] GPT-6 Astra: OpenAI’s biggest LLM launch of all time

new SOTA computer use and coding, 2.5x pricier per token, but WAY cheaper per task, less monitorable. overall, a very successful launch of OpenAI’s new frontier model class.

中文介绍 OpenAI发布GPT-6 Astra,这是其有史以来最大的LLM发布,使用成本更高但任务成本更低。

Daybreak for Frontline Defenders: $1B to protect essential services

OpenAI introduces Daybreak for Frontline Defenders. A $1 billion commitment expands access to frontier cyber AI, training, and support for essential services.

中文介绍 OpenAI推出Daybreak for Frontline Defenders,承诺10亿美元以扩大前沿网络安全AI、培训和必要服务的访问。

Legora reviewed 41 documents in minutes with GPT-6 Astra

Legora used GPT-6 Astra to review 41 documents in minutes, find all four planted errors, and improve performance by nearly 40% in this financial-review workflow.

中文介绍 Legora使用GPT-6 Astra在几分钟内审查41份文件,发现所有四个植入的错误,并在财务审查流程中提高了近40%的性能。

Canada’s retaliatory US tariffs set to take effect as trade dispute grows

Counter-measures covering several sectors come after Trump announced 50% tariffs on Canadian goods Canada is set to impose retaliatory tariffs on billions of dollars’ worth of American imports early on Tuesday, escalating a trade fight with its largest trading partner as tensions between US presiden

Hawaii braces for Hurricane Lowell

Threats of cyclones and deadly surf as the Category 3 storm path approaches Hawaiian islands on Monday night.

Containing Germany’s Far Right

The AfD victory in a state election is forcing a rethink of the country’s postwar strategy to safeguard democracy.

Yen Gains Past 154 as Rally Extends

The yen extends its recent rally, trading beyond 154 per dollar. Bloomberg's Mark Cranfield has the latest. (Source: Bloomberg)

中文摘要 日元汇率持续上涨,突破154日元兑1美元大关。

Gold Steadies as Traders Weigh Mideast Tensions and Dollar Moves

Gold was steady near $4,400 an ounce, as traders weighed the competing effects of Middle East tensions and the US dollar’s decline against the yen.

中文摘要 金价稳定在每盎司4400美元左右,投资者权衡中东紧张局势和美元对日元贬值的影响。

AI cancer cures slowed by chip shortage, says UK's biggest tech boss

The head of chip designer Arm says modelling how a DNA marker is impacted by cancer cannot be done now, but computers are "going to solve it".

中文摘要 芯片短缺导致AI癌症治疗方法研究受阻,英国最大科技公司负责人表示电脑将解决这一问题。

CapitaLand Targets $500 Million for Third Asian Credit Fund

Singapore’s CapitaLand Investment Ltd. is targeting $500 million in commitments from investors for its third Asia Pacific credit program, according to people familiar with the matter.

中文摘要 新加坡的凯德集团计划为其第三个亚太地区信贷项目筹集5亿美元。

Latest Oil Market News and Analysis for Sept. 8

Oil extended gains as traders watched for details of an Iranian deal with Oman to manage shipping through the Strait of Hormuz, which could tighten Tehran’s control over the crucial waterway.

中文摘要 油价上涨,交易者关注伊朗与阿曼关于霍尔木兹海峡航运的协议细节。

Stocks in Asia Poised to Edge Down, Yen in Focus: Markets Wrap

Asian stocks were set to edge lower as tensions in the Middle East pushed oil prices higher, adding to inflation concerns and keeping investors wary about further monetary policy tightening. The yen extended gains to reach its strongest level since February.

中文摘要 亚洲股市预计小幅下跌,日元汇率上涨至2月份以来最高水平。

Sapporo Breweries CSO on Tariffs, Iran War

Sapporo Breweries is shifting some production to the US from Canada due to 50% tariffs imposed on beer exports from the country, Chief Strategy Officer Rieko Shofu said. She spoke exclusively with Bloomberg's Shery Ahn on being impacted by tariffs, as well as the Iran war. (Source: Bloomberg)

中文摘要 札幌啤酒因加拿大啤酒出口关税提高50%,将部分生产转移到美国。

Why Are Japan’s Bond Yields Rising? What’s Driving the Surge

Government bonds are generally viewed as one of the safest assets to invest in. That’s because it’s considered relatively unlikely that the issuer — a government — will go bankrupt. Among the world’s major government debt markets, Japan’s $7.5 trillion bond market has, for decades, been considered o

中文摘要 日本国债收益率上升,分析其背后的驱动因素。

Bathla Group Collapse: What Happens to Homebuyers and Their Deposits?

Sydney property developer Bathla Group’s collapse has left thousands of Australian homebuyers — who have paid deposits and are waiting for their off-the-plan homes to be built — in limbo.

中文摘要 悉尼开发商Bathla集团破产,数千名购房者处于困境。

Investigators begin work on why cargo plane overran Miami runway

Five people died and five others were injured as the Amazon plane overshot the runway after landing at Miami International Airport.

中文摘要 亚马逊货运飞机在迈阿密国际机场降落时冲出跑道,造成5人死亡,5人受伤。

该源今日无内容。