tester-army/e2e
TypeScript · ★ 6,358 · 🍴 285 · 📈 1,725 stars today
Next generation e2e testing framework for web and mobile apps.
中文介绍 提供下一代端到端测试框架,适用于网页和移动应用,用于提升测试效率和稳定性。
TypeScript · ★ 6,358 · 🍴 285 · 📈 1,725 stars today
Next generation e2e testing framework for web and mobile apps.
中文介绍 提供下一代端到端测试框架,适用于网页和移动应用,用于提升测试效率和稳定性。
Shell · ★ 278,171 · 🍴 23,291 · 📈 889 stars today
Skills for Real Engineers. Straight from my .agents directory.
中文介绍 为真实工程师提供技能集,源自作者的个人技能目录,可用于提升个人技能管理。
Python · ★ 17,990 · 🍴 1,805 · 📈 619 stars today
Give your agent CAD superpowers.
中文介绍 赋予代理CAD能力,通过文本转换为CAD设计,适用于需要快速生成CAD模型的场景。
C++ · ★ 6,591 · 🍴 489 · 📈 949 stars today
Tool for automatic PS5 executables porting to Linux and Windows
中文介绍 自动将PS5可执行文件移植到Linux和Windows系统,适用于需要跨平台运行PS5游戏或应用的开发者。
JavaScript · ★ 77,707 · 🍴 4,629 · 📈 616 stars today
The design language that makes your AI harness better at design.
中文介绍 设计语言,通过优化AI设计过程,提升设计效率和效果。
TypeScript · ★ 97,193 · 🍴 8,564 · 📈 534 stars today
Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More
中文介绍 为每个代理提供会话间的持久上下文,通过AI压缩和注入上下文,适用于需要持续学习上下文的智能代理。
Python · ★ 54,424 · 🍴 3,125 · 📈 326 stars today
A skill to stop your coding agent from burying the answer. ADHD-friendly output.
中文介绍 专为ADHD患者设计的技能,帮助编码代理避免遗漏答案,提升编码效率。
TypeScript · ★ 9,526 · 🍴 1,059 · 📈 2,956 stars today
Reverse engineer anything with agents, from app behavior down to native binaries.
中文介绍 使用代理进行逆向工程,从应用行为到原生二进制文件,适用于需要逆向分析各种软件的工程师。
Cuda · ★ 8,712 · 🍴 1,372 · 📈 199 stars today
DeepGEMM: clean and efficient BLAS kernel library on GPU
中文介绍 在GPU上提供高效且干净的BLAS内核库,适用于需要高性能数学运算的深度学习应用。
Shell · ★ 157,837 · 🍴 25,448 · 📈 623 stars today
A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables.
中文介绍 提供全面的AI代理服务,包括前端专家、Reddit社区忍者等,适用于需要多样化AI代理的用户。
JavaScript · ★ 5,719 · 🍴 783 · 📈 1,419 stars today
Self-hosted gym & body-weight tracker — plan routines, log workouts (supersets, warm-ups, cardio), see which muscles are trained, fatigued or detrained, import from FitNotes/Strong/Hevy, passkey login. Your data, your server.
中文介绍 自托管健身和体重跟踪器,支持计划训练、记录锻炼、查看肌肉训练状态,适用于健身爱好者。
HTML · ★ 44,048 · 🍴 2,842 · 📈 228 stars today
Editorial diagram design for Claude Code, Codex, GitHub Copilot, Factory Droid, and Pi. 42 diagram types. Self-contained HTML + SVG. No shadows. No Mermaid slop.
中文介绍 为Claude Code、Codex、GitHub Copilot等提供编辑器,支持多种图表类型,适用于需要快速创建图表的设计师。
👍 4
Reinforcement learning (RL) has greatly advanced the capabilities of large language models (LLMs), but its memory demands remain a barrier to broader adoption. We introduce LoGRA, an approach to RL post-training that reduces memory by retaining useful learning signals in low-rank gradient sketches.
👍 2
Memory self-evolution uses task feedback to iteratively improve executable memory programs that store and retrieve information from past interactions. Existing approaches typically adopt holistic evolution, deriving revision directions from mixed feedback and judging progress by overall performance.
👍 1
Tabular foundation models perform in-context learning (ICL) by conditioning predictions on labeled training examples provided as context. Unlike traditional models that separate training from inference, these models must process all training examples in every forward pass, making each prediction exp
👍 1
A language model can give a correct answer more probability than any single incorrect answer and still usually sample an incorrect one, because the incorrect answers together hold more probability. The power distribution raises each complete answer's probability to a power above one and renormalizes
👍 2
Real-world users often exhibit highly heterogeneous preferences over multiple objectives for LLM responses. A lightweight aligner can tailor these responses to individual preferences, but scarce user-specific feedback makes personalized training difficult. Learning shared initializations across user
👍 4
Modern information systems, including many agentic workflows, use dense retrieval to explore large amounts of unstructured data. However, dense retrieval relies on surface-level semantic similarity, which is insufficient for increasingly complex search applications. Here, we investigate agentic retr
👍 2
While deep features have transformed anomaly detection in images and video, their impact on tabular data has been less substantial, partly due to the limited availability of strong deep representations. Recently, prior-data fitted networks (PFNs) have emerged as a promising source of such representa
👍 2
A semantic decision engine such as Jev can return a valid answer and still miss a network deadline, select an infeasible action or leave the service unverified. We systematize 139 paper families by decision interface, execution path and check ownership. Fifty families claim that their engine fits a
👍 6
Latent reasoning lets a large language model (LLM) think in a continuous space and verbalize only the answer. We argue that an effective latent thought must meet five requirements: it should be useful, helping produce the correct answer rather than merely changing it, diverse, so that resampling yie
👍 17
Geometry optimization is a major cost in many quantum-chemical workflows: each optimization step requires one force evaluation, and at the density-functional level that evaluation dominates the wall time. Research in this area has produced a broad range of optimization methods, and we ask whether a
👍 3
Jev-style decision models return categorical probability distributions over predefined options without generating free-form text, enabling software systems to act on their outputs directly. In this work, we investigate the extent to which general-purpose LLMs already possess this capability out of t
👍 8
Tool-using AI agents are increasingly deployed across enterprise software systems, yet widely used benchmarks primarily evaluate nominal task completion, conflating baseline planning competence with operational fault recovery. We introduce UndoBench, a benchmark spanning 36 base workflows and 36 fau
👍 14
Search agents repeatedly make short decisions about relevance, evidence sufficiency, and search actions. Using generative language models for these decisions introduces latency and unreliable confidence. We present SearchJev, a fast and calibrated System-1 model that separates search decisions from
👍 1
Routine complete blood counts (CBCs) could yield new biomarkers, but the private records needed to evaluate candidates cannot be shared with frontier language model agents that excel at discovery. We distilled the evidence held in the Clalit Health Services panel of over 5.4 million patients into a
👍 1
Numerical Linear Algebra (NLA) has consistently played a vital role in advancing science by providing tools to solve fundamental problems encountered in scientific and engineering applications. Over the decades, it has continually evolved to meet the demands driven by successive waves of scientific
👍 24
LLM-based agents are increasingly capable of generating complex 3D structures, with the potential to reshape how objects are designed and realized in the physical world. Yet, producing elegant geometry is fundamentally different from producing objects that can be built and perform their intended fun
👍 4
Real-world time series are frequently driven by exogenous events and structural shifts, rendering conventional forecasting based solely on historical numerical observations insufficient. While language models can retrieve external news, standard retrieval-augmented approaches struggle with high nois
👍 2
Autoregressive transformers remain comparatively weak for protein sequence and structure generation. We study the role of target representation: amino acid tokens encode residue identities without explicit contextual semantics, while backbone coordinates require a discrete representation in our fram
👍 0
Existing streaming vision-language models (VLMs) continuously perceive and reason over visual streams, but their computational pathways remain fixed throughout inference. Consequently, they cannot adapt computation to evolving scene dynamics, where different future events demand different levels and
👍 2
Intent-based Open RAN needs an interpreter that turns intents into A1 policies within the loop of the RAN intelligent controller (RIC). Decision models such as Jev-1.13.0 return typed policy fields, whereas generative large language models (LLMs) produce the policy token by token. We ask whether the
👍 0
Advances in experimental instrumentation and automation generate increasingly rich datasets, but turning experimental observations into microscopic understanding remains a bottleneck in scientific discovery. To accelerate this process, we introduce AI Theorist, a system of artificial intelligence (A
👍 9
Interactive world models are increasingly capable of generating environments and acting within them, yet deliberately editing an existing executable world remains underexplored. We formulate world editing as intervening on an existing world while preserving properties that should remain unchanged, a
👍 7
Computer-use agents need to reliably ground action targets in complex desktop scenes, where multiple applications, overlapping windows, and visually similar controls compete for attention. Existing training data rarely pair such scenes with dense annotations or vary them in a controlled way. We intr
👍 0
Recent years have seen the employment of a plethora of machine learning (ML) models in high-stakes domains, but they remain largely opaque to the practitioners who act on their predictions. While post-hoc explanation methods offer a lens into this model behavior, wielding them effectively demands ex
👍 17
Recent advances have enabled unified omni-modal models in understanding audio, vision, and language. However, existing benchmarks, training data, and learning methods largely treat the modalities independently, leaving the capability of audio-visual joint reasoning poorly evaluated and insufficientl
👍 8
Manipulation behaviors vary widely across objects and scenes, but they share a small set of reusable skills, and planning with these skills helps embodied agents generalize to new tasks. Yet an agent can only plan with skills it knows. Recovering skills from observed experience, the inverse of plann
👍 8
Embodied agents now take on ever longer tasks. For long tasks, knowing only whether a task finally succeeds or fails says little; the steps along the way matter. Progress Reward Models (PRMs) score how far a task has come at every step, and serve as dense rewards, verifiers and monitors. Yet in long
👍 22
Proactive LLM agents can turn idle compute into useful support before users ask. Yet even correct work can misread user context, impose review costs, or undermine trust. This work proposes foundations for designing, realizing, and evaluating proactive LLM agents around three joint principles (3T): T
👍 8
Does conversational memory need LLM-extracted facts, or is selecting the right raw turns enough? Published results disagree. Extraction-based systems report gains from distilled facts. Recent studies find raw history with good ranking does as well, but disagree about whether ranking matters. We ran
👍 3
Face Image Quality Assessment determines the suitability of captured face images for automated face recognition (FR), a critical capability for reliable biometric systems. Existing state-of-the-art FR-integrated FIQA methods suffer from temporal instability: as the feature space evolves during train
@BrianRoemmele · 489.9K 粉丝 · 1.1M 阅 · 533 赞 · 85 转
You do not build a Grok Bot. You hire one. The difference is the whole article. A chatbot answers a question and forgets the room. A Grok Bot has a name, a job, a conversation that persists, and a
中文介绍 探讨如何雇佣Grok Bot,区别于普通聊天机器人,Grok Bot拥有持续对话能力。
@stevenmarkryan · 458.8K 粉丝 · 124.9K 阅 · 541 赞 · 84 转
Incredible things are happening in AI. What’s more incredible is that it’s still EXTREMELY early. As in, Day 0.000. Things are only going to get crazier from here. In case you’ve been living under a
中文介绍 AI发展初期,预测未来将更加疯狂,呼吁关注AI的早期阶段。
@TryLiveAvatar · 1.1K 粉丝 · 123.4K 阅 · 692 赞 · 188 转
The best way to learn a language was always a patient tutor. For the first time, everyone can have one. Here's what that changes for learners, teachers and schools. The hardest part is speaking Ask
中文介绍 语言学习将迎来变革,AI导师让每个人都能拥有个性化的学习体验。
@mustafasuleyman · 1.2M 粉丝 · 99.2K 阅 · 587 赞 · 104 转
This is the prediction Nobel laureate Daron Acemoglu makes in the first issue of The Humanist Review, our new magazine exploring the future of AI, published by MAI. He argues we need to stop
中文介绍 诺贝尔奖得主预测,10年内AI将取代人类工作的比例仅为5%。
@gregisenberg · 724.2K 粉丝 · 44.1K 阅 · 569 赞 · 53 转
For the FIRST TIME in HISTORY, you can describe a physical product in one sentence and, a week later, hold it in your hands, cut from metal and built to your exact measurements. That's VIBE
中文介绍 VIBE制造技术突破,首次实现一周内将金属产品设计成实物。
@onlyonealexia · 1.6K 粉丝 · 26.5K 阅 · 509 赞 · 53 转
A List of Web3, AI & Web2 Hackathons Compiled by @onlyonealexia There are a lot of hackathon lists that look impressive until you actually open the links. I compiled a list of ongoing and upcoming
中文介绍 分享Web3、AI和Web2的Hackathon列表,筛选出值得参与的竞赛。
@Hesamation · 90.6K 粉丝 · 21.4K 阅 · 618 赞 · 51 转
if you have an ADHD brain, you probably have six different goals, you get overwhelmed, and you miss days without working on any one of them. that's me. I love agents, harnesses, RL, LLM architecture,
中文介绍 针对多兴趣人士,介绍如何构建AI系统来管理多重目标。
Atlassian and OpenAI are expanding their partnership to connect frontier models with enterprise knowledge and help teams plan, build, and deliver work.
中文介绍 Atlassian与OpenAI扩大合作,将前沿模型与企业知识连接,助力团队规划、建设和交付工作。
Jump Trading uses OpenAI to expand quantitative research. See how longer-running AI workflows combine multiple data sources with human review.
中文介绍 Jump Trading利用OpenAI扩展定量研究,通过结合多个数据源和人工审核的AI工作流程。
OpenAI publishes new results on open problems in mathematics from an internal frontier model and shares Lean proof formalizations and research details on GitHub.
中文介绍 OpenAI发布关于数学开放问题的内部前沿模型新结果,并在GitHub上分享Lean证明形式化和研究细节。
Learn how OpenAI and Ironclad are training and evaluating AI agents on complex contracting workflows to advance computer use for professional work.
中文介绍 OpenAI与Ironclad合作,通过训练和评估AI代理在复杂合同工作流程上的应用,推进计算机在专业工作中的应用。
中文介绍 Falcon-Emirati:当一个LLM学习方言、文化和细微差别。
A small win for US open source
中文介绍 美国开源模型[AINews]Reflection Beam - 501B-A23B取得小胜。
中文介绍 本能小组讨论,Reflection的501B模型,OpenAI文本水印。
For all the data that AI systems continually amass and analyze, enterprise AI agents often suffer from a curious shortcoming: a lack of knowledge. More than data, knowledge is the understanding of what the data means in the context of individual organizations. AI agents need this understanding to re
中文介绍 将AI代理连接到企业知识。企业AI代理往往缺乏知识,需要理解数据在特定组织背景中的意义。
How OpenAI is approaching text watermarking under EU rules. Learn where watermarks apply, how detection works, and why access starts with researchers.
中文介绍 OpenAI阐述遵守欧盟文本溯源规则的方法,介绍水印应用、检测方式及为何研究者优先获取访问权。
In 2026, the question for enterprise AI is no longer whether predictive models can outperform statistical forecasts—that argument is settled. The big question now is how to enable predictive systems to act on their own conclusions without drifting from business intent. The frontier has moved from pr
中文介绍 将预测分析引入代理AI时代。企业AI的核心问题是使预测系统能够根据自身结论行动,而不偏离商业意图。
OpenAI introduces a new visual ad format in ChatGPT and expands measurement tools, attribution partnerships, and brand suitability for advertisers.
中文介绍 OpenAI在ChatGPT中推出新的视觉广告格式,并扩大测量工具、归因合作伙伴关系和品牌适用性。
Over the summer I talked to the CEO of Springboards, a startup building an LLM that’s designed to come up with a wider variety of responses than its mainstream rivals do. At the start of the call, he said something that’s been stuck in my head since: “We often say that we’re a self-loathing AI…
中文介绍 人们真的很讨厌AI,但他们为何不能获得足够的应用?
Yossi Matias, Vice President & Head of Google Research, explores how AI is beginning to reshape biology, infrastructure, manufacturing, and science, and why its greatest impact may come when it intersects with other fields. Step inside the newsroom with our MIT Technology Review editors for sharp an
中文介绍 EmTech Future 2026:当AI遇到一切。Yossi Matias探讨了AI如何开始重塑生物、基础设施、制造和科学,以及为什么它最大的影响可能发生在与其他领域的交叉处。
中文介绍 Prime Inference、代理人群、Claude学院。
中文介绍 代理说已完成,但数据库不同意。
Ruling on Mount Pleasant coalmine shows ‘we cannot continue to dig up coal … and pretend the consequences have nothing to do with us’, group says A Hunter Valley community group has won Australia’s first high court case to consider climate change, in a ruling advocates say sets a binding national pr
Experts and lawyers called the apparent recovery unprecedented, noting she is believed to be the only known person to survive a lethal injection.
The mountainous district is key to control of Mocha, the Red Sea port the Houthis seized last month.
Australian adults again able to access the world’s most popular porn site after parent company introduces age checking via Apple devices Get our breaking news email, free app or daily news podcast Australian adults will be able to access Pornhub once again after the site brought in age-checking on A
At least 10 civilians were killed in West Kordofan and 5,500 displaced in Blue Nile as clashes widened on Tuesday.
More than one million people could lose life-saving aid by October as funding shrinks and arrivals from Sudan continue.
Saint-Denis Mayor Bally Bagayoko was tear-gassed after shielding student protesters from riot police.
Yemeni forces claim they have tightened control over the highest peak in the Jabal Habashi district of Taiz.
Commissioner said Shenzhen Qingcheng, the app maker behind the HeyCyan app in the smartglasses, failed to respond to her inquiries Get our breaking news email, free app or daily news podcast The Australian privacy regulator has opened an investigation into the China-based software company behind the
Turkiye’s first domestically produced electric SUV, the Togg T10X burst into flames at a charging station.
World Health Organization detects no sign of further outbreak since Darya Shipilova died last week The World Health Organization has said the risk of an epidemic in Russia is low after the sudden death of a lab technician who worked at a plague research institute in Siberia. Darya Shipilova, who was
Government report finds the country could be hit with more frequent and intense heatwaves, as well as damaging storms and cyclones New Zealand could be hit with more frequent and intense heatwaves and catastrophic glacier melts in the near future, as greenhouse gas levels hit record highs, a new rep
The Bolsonaro family emerged triumphant in Brazil’s elections, mounting a surprising political resurrection in Latin America’s largest nation.
A sharp increase in human activity is affecting the way the Himalayan monal lives and communicates, studies show.
Frequent outages have fuelled protests against the government of interim Venezuelan President Delcy Rodriguez.
Taiwan is emerging as the stronger bet for equity investors due to its deeper linkages across the AI supply chain and a more upbeat earnings outlook. Bloomberg's Winnie Hsu breaks down the news. (Source: Bloomberg)
中文摘要 台湾因AI供应链联系更紧密和乐观的盈利预期,成为股民更看好的投资地,超越韩国成为全球市场领导者。
Former Canadian Minister of International Trade and Milken Institute Senior Fellow Mary Ng outlines Canada's expanding trade push across Asia and the Indo-Pacific, emphasizing that deeper market access in energy, technology, and agriculture could reach 3 billion consumers and drive export growth. (S
中文摘要 加拿大前国际贸易部长玛丽·吴强调,加拿大在亚洲及印太地区的贸易扩张,有望触及30亿消费者,促进出口增长。
As equity investors enter the final stretch of a pivotal year that’s propelled South Korea and Taiwan to the forefront of the global AI trade, the latter is emerging as the stronger bet.
Two years after JLR relaunched Jaguar into a blizzard of controversy, it has unveiled its new electric car.
中文摘要 两年后,捷豹路虎在争议中重启了捷豹品牌,并推出了其新款电动汽车。
Asset Management One Co. plans to boost spending on talent over the next three years as rising Japanese interest rates broaden investment choices and push pension funds, university endowments and individuals to rethink how they allocate their money.
中文摘要 资产管理公司One计划在未来三年内将人才支出提高30%,以应对日本利率上升带来的投资选择增加。
British carmaker unveils £130,000 electric car following radical rebrand
中文摘要 英国汽车制造商推出了一款价值13万英镑的电动汽车,这是其品牌重塑后的新产品。
Gold held gains as increasing oil supplies from the Middle East and a decline in bond yields eased pressure on the US Federal Reserve to hike interest rates this month.
中文摘要 随着中东地区石油供应增加和债券收益率下降,黄金价格保持稳定,缓解了美联储本月加息的压力。
BHP Group will sell its Kambalda nickel concentrator plant and associated land in Western Australia to Gold Fields Ltd. for an undisclosed price.
中文摘要 必和必拓公司将位于澳大利亚的Kambalda镍精炼厂及其相关土地出售给Gold Fields Ltd.,交易价格未公开。
Samsung Electronics Co.’s earnings will test whether it can convince investors of its long-term outlook, an increasingly critical task as record profits have failed to reinvigorate the stock.
中文摘要 三星电子的收益将考验其能否让投资者相信其长期前景,因为创纪录的利润未能提振股价。
Ken Leech, the former co-chief investment officer at Western Asset Management Co., agreed to pay $3 million to settle a US Securities and Exchange Commission lawsuit claiming he’d cherry picked winning trades.
中文摘要 西部资产管理公司前联合首席投资官肯·利奇同意支付300万美元以解决美国证券交易委员会对其选择性交易的诉讼。
Frozen yoghurt, which was a huge craze in the 2000s and 2010s, has made its return with multiple chains popping up across the wider UK.
中文摘要 冰沙卷土重来,但每桶12英镑的价格能维持多久?
Order of the Sinking Star is due out on Thursday and has roughly 1,500 puzzles for players to try
中文摘要 《沉没之星》将于周四发布,包含大约1500个谜题供玩家尝试。
13 回复 · Apple 节点
24 回复 · Python 节点
89 回复 · 程序员 节点
9 回复 · 程序员 节点
6 回复 · 程序员 节点
5 回复 · 程序员 节点
59 回复 · 程序员 节点
11 回复 · 程序员 节点
10 回复 · 程序员 节点
17 回复 · Apple 节点
该源今日无内容。