paperclipai/paperclip
TypeScript · ★ 89,635 · 🍴 15,636 · 📈 2,527 stars today
The open-source app everyone uses to manage agents at work
中文介绍 Paperclip 是一款开源应用,用于管理工作中的智能代理,简化代理管理流程。
TypeScript · ★ 89,635 · 🍴 15,636 · 📈 2,527 stars today
The open-source app everyone uses to manage agents at work
中文介绍 Paperclip 是一款开源应用,用于管理工作中的智能代理,简化代理管理流程。
Python · ★ 37,121 · 🍴 4,825 · 📈 4,463 stars today
Hindsight: Agent Memory That Learns
中文介绍 Hindsight 是一款智能代理记忆学习工具,帮助代理从经验中学习。
Python · ★ 39,877 · 🍴 4,740 · 📈 3,060 stars today
VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.
中文介绍 VoiceStudio 是一款开源的语音克隆和设计工具,支持多种语言的视频配音、语音识别、转录和有声书创作。
Python · ★ 59,190 · 🍴 10,222 · 📈 848 stars today
Learn it. Build it. Ship it for others.
中文介绍 AI Engineering From Scratch 是一个学习人工智能工程实践的项目,旨在帮助用户从零开始构建和部署人工智能应用。
Shell · ★ 6,539 · 🍴 219 · 📈 139 stars today
An open-source Android app to let you browse YouTube and other services freely.
中文介绍 PipePipe 是一款开源的 Android 应用,允许用户自由浏览 YouTube 和其他服务。
TypeScript · ★ 5,362 · 🍴 138 · 📈 186 stars today
TypeScript-to-Native Compiler
中文介绍 scriptc 是一个 TypeScript 到本地编译器的工具,用于将 TypeScript 代码编译成原生应用。
TypeScript · ★ 900 · 🍴 97 · 📈 114 stars today
Multi-agent harness that runs Claude Code and Codex together as one system
中文介绍 openrig 是一个多智能代理工具,可以将 Claude Code 和 Codex 结合运行,形成一个统一的系统。
TypeScript · ★ 20,113 · 🍴 1,703 · 📈 920 stars today
The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.
中文介绍 Univer 是一个适用于人工智能代理的办公套件,支持电子表格、文档、幻灯片、画布、关系表和 PDF 等功能。
C · ★ 786 · 🍴 152 · 📈 117 stars today
Run x86-64 Windows PC games on jailed iOS via FEX-Emu + Wine + DXMT
中文介绍 Madeira 是一个运行在 iOS 设备上的 x86-64 Windows PC 游戏模拟器,利用 FEX-Emu、Wine 和 DXMT 技术。
👍 8
In this paper, we propose RGBD20K, a novel dataset for facilitating the development of more robust and general RGB-D semantic segmentation by encompassing abundant categories and high-quality annotations. RGBD20K possesses several attractive properties: (1) Expanded Semantic Space. In particular, it
👍 75
While Large Language Models (LLMs) rely on highly non-linear components, in this work we demonstrate that they exhibit fundamental linearity: when inputs from distinct text streams are linearly combined, the model outputs a superposition of the individual next-token distributions. We term this the S
👍 7
Detectors of alignment failures screen deployed language models and score alignment benchmarks. Most are generative judges that spend a decoding pass on every criterion, and classifiers that read token probabilities, such as Llama Guard, still score one fixed label per call. Jev, a model trained wit
👍 11
Task and motion planning (TAMP) problems remain difficult even with full observability and object-centric states because discrete decisions are tightly coupled to geometric, kinematic, and dynamic constraints. Generalized TAMP addresses this difficulty by exploiting regularities across problem insta
👍 17
Rufus-Air is an open and reproducible post-training recipe on GLM-4.5-Air-Base (106B-A12B), organized as a serial pipeline of eight stages: SFT, Reasoning RL, Coding RL, Instruction-Following RL, General Agent, Coding Agent, Search Agent, and RLHF. We document the data, reward design, infrastructure
👍 9
Deep search requires LLM agents to decompose complex queries, search for evidence, and synthesize grounded answers, yet existing ReAct-style agents suffer from two limitations: role coupling, where one policy must handle planning, evidence use, and synthesis; and context accumulation, where growing
👍 6
We introduce PUBG Ally, an embodied agent for PUBG: BATTLEGROUNDS that can reason, act autonomously, and play alongside players as a voice-enabled teammate. Building such a teammate requires combining two difficult capabilities: it must perceive and respond to a constantly changing game world under
👍 7
General-purpose vision-language models (VLMs) bring broad knowledge and spatial reasoning to robot manipulation, yet existing systems either use them indirectly, to predict constraints or write programs, or give them a view of the scene rather than a world in which to act. We present World Action Ag
👍 2
Block-based video codecs select coding parameters based on the input by optimizing a rate-distortion trade-off. The conventional distortion choice, the sum of squared errors (SSE), simplifies parameter selection: the SSE is the sum of block-wise SSEs, so rate-distortion optimization (RDO) can treat
👍 11
The rapid progression of large language models is extending AI from passive content generation into the active workflows of engineering and scientific discovery. This shift raises a compelling question: can AI be both the object of development and an active participant in building next-generation AI
👍 11
Recently, Large Language Models (LLMs) have been increasingly able to solve advanced mathematical problems, including many that have been open for decades. This opens the door to expansion of mathematical knowledge at unprecedented scale. Yet, while LLMs may be able to conjecture and prove more and
👍 4
World-action models (WAMs) transfer visual and motion priors from pretrained video generators to robot control by jointly modeling visual dynamics and actions. Existing WAMs, however, predict dense future frames during training, repeatedly modeling largely unchanged content and coupling action-condi
👍 203
Object permanence and solidity are hallmarks of human cognitive priors. Recent studies show that video generation models, a paradigmatic class of current world models, have begun to show emerged reasoning abilities, making them ideal candidates for building human-like physical intelligence. Do video
👍 19
Recent advances in large language models (LLMs) have enabled agents to tackle long-horizon tasks across diverse environments. To further improve agent performance, existing language world models typically predict environment observations, yet reconstructing high-entropy, execution-dependent tool res
👍 4
Solutions based on large language models (LLMs) often rely on temperature sampling to improve accuracy and stability by aggregating multiple samples from the completion distribution. However, this memoryless approach is inherently suboptimal: because it lacks awareness of prior generations and their
👍 35
Agentic memory systems reuse past experience to improve future performance, yet most existing designs curate memory at write time: once a task is completed, its trajectory is distilled into a fixed artifact, such as a reflection, workflow, skill, or reasoning strategy, that is later retrieved by sim
👍 3
We study coding agents for long-horizon, dexterous robotics and ask whether their solutions can provide scalable supervision for learning general robot policies. To test this, we develop EMBODIEDSWE-BENCH, a simulation benchmark for coding agents spanning contact-rich manipulation, deformable object
👍 3
Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at the frontier has focused primarily on advancing model capabilities. We introduce StudentBench, a suite of AI teaching evaluations and a public platform that enables large-scale data collection
👍 2
We introduce Knowledge Pull Requests (KPRs), a framework for continual document authoring that makes each change interpretable. Documents require ongoing revision as new knowledge surfaces from other sources, languages, or times, but existing approaches either edit with no account of what knowledge
👍 15
The rapid capability gains of frontier language models are widely attributed to improved reasoning abilities, yet this cannot be verified as raw CoT traces in closed-source systems are hidden. By registering a simple custom tool through a standard API feature, we induce frontier models to externaliz
👍 2
Calibration of language models -- the alignment between expressed or implicit confidence and empirical correctness -- is a well-studied subfield within NLP. Methods to measure it already exist. The problem is adoption: outside this subfield, NLP research regularly introduces new models, datasets, an
👍 5
Learned-memory methods store information in an explicit table and consume it through a separate reader, allowing addressing, storage, and reading to be modified independently. We study whether useful memory can also be generated rather than only retrieved. MemoryAthena uses three pathways: direct En
👍 2
Task planning bridges high-level instructions and executable behavior in long-horizon manipulation, yet modern Vision-Language-Action (VLA) systems often leave this intermediate structure implicit. Existing chain-of-thought (CoT) planners also tend to rely on coarse task-level annotations or seriali
👍 43
Evaluating world models requires assessing both the quality of the worlds they generate and their consistency and responsiveness under exploration, interaction, and modification. We introduce HappyWorld-Bench, a comprehensive benchmark that evaluates whether generated worlds remain reliable as agent
👍 22
Humans can effortlessly localize the direction of a sound source and integrate it with visual cues for reasoning, yet this remains challenging for embodied agents. In particular, it is still unclear how to effectively evaluate and model spatial audio understanding in embodied settings. To address th
👍 11
Robotic bin packing requires long-horizon sequential decision-making, as each object placement affects the available space for subsequent packing. Existing methods primarily rely on hand-crafted geometric heuristics that optimize predefined objectives or reinforcement learning policies learned throu
👍 8
Modern Transformer design and compression both reduce to allocating capacity under a budget. The standard scalars for these decisions, #Params and #FLOPs, capture size and compute but not architectural structure: two architectures with identical parameter budgets but different depth-width, head, or
👍 6
Collective intelligence depends not only on what team members know, but also on how they organize their work. When the structure of a solution is unknown, useful roles and divisions of labor cannot be specified in advance; teams must learn from experience how to organize reasoning as it unfolds. Hum
👍 48
Spatial reasoning is essential for vision-language models (VLMs) to understand and act in the physical world. Reasoning in dynamic environments requires VLMs to perceive local state transitions caused by object motion and viewpoint changes and integrate them over long trajectories to maintain an upd
👍 7
Modern AI agents routinely cross trust boundaries: they ingest untrusted content, combine it with privileged instructions, persist intermediate beliefs in long-term memory, and invoke privileged tools. This creates an attack surface in which malicious payloads can enter through model inputs and caus
@jacobandreou · 7.8K 粉丝 · 521.0K 阅 · 500 赞 · 77 转
The bottleneck is no longer intelligence. Over the last six months, intelligence has accelerated dramatically. But smarter models are not enough for everyone to benefit from novel intelligence. New
中文介绍 探讨AI智能瓶颈,强调模型智能提升对大众受益的重要性。
@DhravyaShah · 63.2K 粉丝 · 112.4K 阅 · 608 赞 · 45 转
It's been just 2 weeks since we discontinued the company brain harness. We had to offboard tons of people to other products, but many of our customers, and others suggested us to open source our work
中文介绍 分享开源公司脑力工具设计过程,回顾产品转型过程。
@chamath · 2.4M 粉丝 · 97.8K 阅 · 551 赞 · 44 转
In June, I wrote that vibe coding was dead and that ROI-driven analysis of AI was about to go from a nice-to-have to a necessity. Here's what AI ROI is and why it's tricky to see in GDP numbers so
中文介绍 分析AI繁荣与生产力提升的差距,探讨AI投资回报率在GDP中的体现。
@saylor · 5.2M 粉丝 · 96.2K 阅 · 661 赞 · 98 转
Artificial intelligence will make it possible for individuals and companies to produce far more than they can today. That makes the freedom to create, finance, own, and exchange things more important.
中文介绍 展望数字经济发展,强调创造、融资、拥有和交换的自由重要性。
@0xMovez · 36.4K 粉丝 · 78.9K 阅 · 737 赞 · 46 转
Most people who try motion design with Opus 5.5 end up with the same video: centered text on a gradient, everything fading in, a logo at the end. They don't give it a reference, don't give it a
中文介绍 分享使用Opus 5.5构建动态设计工作室的教程,避免常见设计陷阱。
@trq212 · 354.7K 粉丝 · 62.7K 阅 · 965 赞 · 52 转
One of the best parts of our newest Claude models is how they respond to effort without breaking the prompt cache in Claude Code, but I’ve received a lot of questions on this from users. What is
中文介绍 解析Claude Code中模型对用户努力的反应,解答用户疑问。
@cerebras · 76.6K 粉丝 · 37.6K 阅 · 554 赞 · 28 转
Written by @milksandmatcha and @0xSero A personal assistant should save you time and effort. Over the past few weeks, we’ve been obsessively testing AI personal assistants on everyday tasks, from
中文介绍 介绍20倍速度的Grok机器人,测试AI个人助理在日常工作中的表现。
In 2023 most people doubted that there could be more than 1 or 2 frontier model labs. Now there are dozens.... and Stripe just bought the best known one for $7B.
中文介绍 Stripe以70亿美元收购了前沿模型实验室中最为知名的一家,而2023年大多数人还怀疑是否可能存在超过1或2家这样的实验室。
With Codex, GPT-Live-1, and GPT-6 Astra, Proaction builds, operates, and sells modern fleet management faster.
中文介绍 Proaction利用Codex、GPT-Live-1和GPT-6 Astra,使现代车队管理更快地建设、运营和销售,从而提升了60%的销售额并节省了75小时以上。
The US government wants to spend $30.3 million over the next five years on an improved form of lie detector, according to a Department of Defense budget request. The program, called Polygraph+ or Polygraph Next, will focus on scoring algorithms that use artificial intelligence and machine learning a
中文介绍 美国国防部请求在未来五年内花费3030万美元研发一种改进型测谎仪,名为Polygraph+或Polygraph Next,将专注于使用人工智能的评分算法。
A quiet day lets us discuss the work behind the scenes - now open for business!
中文介绍 Latent Space透露了幕后的工作现在对外开放。
GWM Worlds 2 uses persistent context and timed actions to steer a world model generating video and audio in real time.
中文介绍 GWM Worlds 2利用持续上下文和定时动作来引导生成视频和音频的世界模型,实现实时世界工程。
中文介绍 ChatGPT Pro Max、Muse实时头像和DeepSeek 10亿美元的年度经常性收入。
Guest Post: In science, thinking has gotten cheap but doing has not. This asymmetry is reshaping how research companies operate, largely inconspicuously.
中文介绍 科学领域思考成本降低,但实际操作成本未减,这种不对称性正在改变研究公司的运营方式。
中文介绍 通过LFM2.5-VL-DSpark加速视觉语言模型的加速。
Team Zuck is absolutely on fire.
中文介绍 Meta Connect 2026大会上,Zuck团队展示了Muse眼镜、语音、视频和Charm技术。
中文介绍 Gemini TTS、Claude的新酶和Google的私有记忆。
中文介绍 介绍如何使用NVIDIA Warp和MjWarp来加速机器人仿真和学习工作流程。
Marking two years of OpenAI Academy and bringing AI skills to even more communities.
中文介绍 OpenAI庆祝OpenAI学院成立两周年,致力于将AI技能带到更多社区。
Radical Numerics is using biological chain-of-thought and multimodal perception to keep up with the bio-defense arms race, design new genomes and gain insights into biology itself.
中文介绍 Radical Numerics利用生物思维链和多模态感知来跟上生物防御军备竞赛,设计新的基因组并深入了解生物学本身。
OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.
中文介绍 OpenAI将其Daybreak项目扩展到乌克兰政府,以支持民用基础设施的网络安全。
OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.
中文介绍 OpenAI首席执行官Sam Altman在联合国安理会就AI安全、人类控制和国际合作发表讲话。
Follow the day’s news live Get our breaking news email, free app or daily news podcast Gallagher says final budget outcome will show $6bn improvement Katy Gallagher and the treasurer, Jim Chalmers, are set to present the final budget outcome today, which she said will show about a $6bn improvement s
中文摘要 澳大利亚财政部长吉姆·查尔默斯和凯特·加拉格尔将公布最终预算结果,预计将显示约60亿澳元的改善。
Abbas Araghchi, Iran’s foreign minister, said a history of contradictory messaging meant that President Trump’s rejection might not be a final position.
中文摘要 伊朗外长阿卜杜拉希姆·阿拉格奇表示,尽管特朗普总统拒绝停火提议,但可能并非最终立场。
Aleksandar Vucic, one of Europe’s longest-serving leaders, resigned as president, paving the way for him to become prime minister after elections next month.
中文摘要 塞尔维亚长期领导人亚历山大·武契奇辞职,为下月成为总理铺平道路。
Student protests at an Indian university have turned violent following claims that a female student was raped on campus.
中文摘要 印度一所大学发生涉嫌性侵事件,导致学生暴力抗议。
U.K. police say five men were arrested on suspicion of preparation of a terrorist act after three vehicles were spotted near an air base the U.S. military has been using to launch attacks on Iran.
中文摘要 五名男子在英国一个被美军用作对伊朗发动空袭的空军基地附近被捕,涉嫌准备恐怖主义行动。
The U.S. military formally ends a more than two-decade-long presence in Iraq this week.
中文摘要 美国军队准备结束在伊拉克超过二十年的存在。
Circulating footage showed significant flooding, following heavy rain in southeastern Algeria.
中文摘要 阿尔及利亚东南部发生严重洪水,道路被淹。
Goncalo Ramos starts in place of Cristiano Ronaldo and scores the winner for Portugal in Oslo.
中文摘要 贡卡洛·拉莫斯取代克里斯蒂亚诺·罗纳尔多,帮助葡萄牙在挪威以2-1战胜卫冕冠军。
A strike on a market in the Yemeni city of Taiz has killed at least seven people and wounded 40.
中文摘要 也门塔伊兹市的一个市场遭到空袭,造成至少7人死亡,40人受伤。
In rare access to Yemen's conflict zone the BBC travels to the front line with pro-government soldiers.
中文摘要 BBC报道了也门冲突前线的情况,与政府军士兵一同前往前线。
An Israeli Apache helicopter struck a commercial centre near Mayfadoun in Southern Lebanon.
中文摘要 以色列在南部黎巴嫩继续空袭,尽管有停火协议。
In an interview with The New York Times, South Korea’s president dropped a bold proposal for how to curb North Korea’s nuclear program.
中文摘要 韩国总统提出一项大胆计划,以遏制朝鲜的核计划。
Five men have been arrested near RAF Fairford in the UK on suspicion of explosives and terrorism offences.
中文摘要 五名男子在英国皇家空军基地附近被捕,涉嫌爆炸和恐怖主义罪行。
Ethiopia’s army has reportedly recaptured the strategic town of Sekota from Tigrayan forces as fighting spreads.
中文摘要 埃塞俄比亚军队据报道重新夺回战略城镇塞科塔,随着冲突蔓延。
Police say the US shooting followed an altercation at an after-hours venue, and four people are in hospital.
中文摘要 底特律一家脱衣舞俱乐部发生枪击事件,造成三人死亡,四人受伤。
The bond market is on the brink of signaling that a series of Federal Reserve interest-rate hikes will start shifting the narrative toward the risk that the US economy stalls out.
中文摘要 债券市场接近发出经济停滞警报,预计美联储连续加息将转向风险。
US businesses far beyond Silicon Valley are adopting Chinese alternatives to OpenAI and Anthropic’s systems
中文摘要 美国企业广泛采用中国AI模型替代OpenAI和Anthropic的系统。
Owners should be free to spend their own money on football — as long as they’re fit and proper
中文摘要 所有者应有权自由支配自己的足球资金,只要他们合格且适当。
The news doesn’t stop when markets close. Hosts David Gura, Christina Ruffini and Lisa Mateo bring clarity, context and a bit of humor to the weekend’s biggest headlines, LIVE from New York. Joined by IAEA Director General Rafael Grossi, The New York Times Finance Reporter Maureen Farrell, Cuban Dep
中文摘要 纽约时间周末新闻节目中,主持人讨论了市场关闭后的新闻,包括IAEA总干事和纽约时报财经记者。
States are feeling more pressure from companies demanding subsidies and tax breaks
中文摘要 各州面临来自要求补贴和税收优惠的公司的更大压力。
Bloomberg entertainment reporter Lucas Shaw is on Bloomberg This Weekend examining Paramount Skydance’s acquisition of Warner Bros. Discovery and the challenge of turning two legacy media companies into a growing business as cable declines and streaming growth slows. Speaking with Lisa Mateo, Shaw s
中文摘要 帕拉蒙天空舞影业收购华纳探索公司,面临将两家传统媒体公司转型为增长业务的挑战。
Bloomberg News Washington breaking news Editor Bryan Pietsch is on Bloomberg This Weekend discussing uncertainty surrounding the Kennedy Center’s renovation, possible demolition and the displacement of major performing arts groups. He tells hosts David Gura and Christina Ruffini that musicians and o
中文摘要 肯尼迪中心音乐家在讨论翻新、可能的拆除和主要表演艺术团体搬迁的不确定性。
Bloomberg White House correspondent and national security editor Michelle Jamrisko is on Bloomberg This Weekend discussing renewed US-Iran diplomacy after indirect talks in New York failed to overcome disagreements over sanctions, the US blockade and the Strait of Hormuz. Speaking with hosts David G
中文摘要 美国-伊朗会谈在纽约间接谈判失败后,由于制裁、美国封锁和霍尔木兹海峡的分歧,讨论了恢复美伊外交。
Wild gyrations in individual stocks on a punchy cocktail of AI euphoria and fear, a tumbling bond market and geopolitical drama bode well for a long-favored trade among hedge fund managers.
中文摘要 人工智能和石油的波动为对冲基金经理中流行的分散交易带来利好。
Renaissance Macro Research economist Neil Dutta tells Bloomberg This Weekend that the US labor market has stabilized while persistent inflation could force the Federal Reserve to raise interest rates at a faster pace than investors currently expect. Speaking with hosts David Gura and Christina Ruffi
中文摘要 美国劳动力市场稳定,持续的通胀可能迫使美联储以比投资者目前预期的更快的速度提高利率。
New York Times finance reporter Maureen Farrell tells Bloomberg This Weekend that Wall Street is growing more skeptical of the AI data center boom as rising interest rates, local opposition and permitting delays add uncertainty and costs to projects. Speaking with hosts David Gura and Christina Ruff
中文摘要 华尔街对AI数据中心热潮日益持怀疑态度,随着利率上升、当地反对和许可延误,项目的不确定性和成本增加。
Artificial intelligence could wipe us all out. Or it might just kill our unwanted subscriptions.
12 回复 · Apple 节点
7 回复 · 程序员 节点
18 回复 · 程序员 节点
23 回复 · 程序员 节点
46 回复 · 程序员 节点
9 回复 · Apple 节点
14 回复 · Apple 节点
11 回复 · Apple 节点
14 回复 · Apple 节点
80 回复 · Linux 节点
该源今日无内容。