openai/codex
Rust · ★ 115,126 · 🍴 17,556 · 📈 2,729 stars today
Lightweight coding agent that runs in your terminal
中文介绍 轻量级终端代码执行代理,用于实现代码的即时执行和调试。
Rust · ★ 115,126 · 🍴 17,556 · 📈 2,729 stars today
Lightweight coding agent that runs in your terminal
中文介绍 轻量级终端代码执行代理,用于实现代码的即时执行和调试。
JavaScript · ★ 12,680 · 🍴 1,427 · 📈 440 stars today
Prompt as Code | GPT-Image2 工业级提示词引擎与模板库,470+ 个案例逆向工程,20+ 套工业级模板,并提炼出Skills,持续更新中
中文介绍 GPT-Image2 工业级提示词引擎与模板库,提供丰富的案例和模板,用于构建图像生成应用。
Shell · ★ 233,809 · 🍴 19,942 · 📈 2,448 stars today
Skills for Real Engineers. Straight from my .agents directory.
中文介绍 为真实工程师设计的技能库,从个人代理目录中提取技能。
Shell · ★ 29,099 · 🍴 2,959 · 📈 814 stars today
Beautiful, Modern & Opinionated Linux
中文介绍 美观且现代的 Linux 发行版,具有独特的风格。
Rust · ★ 14,898 · 🍴 398 · 📈 1,008 stars today
⚡️A native, local-first alternative to Logitech Options+, written in Rust 🦀 — remap buttons, DPI, and SmartShift over HID++. No account, no telemetry.
中文介绍 Logitech Options+ 的本地化替代品,使用 Rust 编写,支持按键重映射、DPI 调整和 SmartShift 功能。
Rust · ★ 30,090 · 🍴 3,825 · 📈 349 stars today
A hive mind communication platform
中文介绍 一个基于蜂群思维的沟通平台,用于集体智慧和协作。
TypeScript · ★ 2,336 · 🍴 269 · 📈 49 stars today
Apache Maka (Incubating) is a local-first AI agent workspace. Model messages, tool calls, tool results, permission decisions, and termination events are recorded as an append-only log.
中文介绍 本地优先的 AI 代理工作空间,记录模型消息、工具调用、结果、权限决策和终止事件。
Python · ★ 47,935 · 🍴 7,887 · 📈 1,040 stars today
Use Claude Code, Codex, Pi, and OpenCode for free (1.3B+ free tokens) from your terminal, app, IDE, or phone like OpenClaw (voice supported + ToS friendly)
中文介绍 从终端、应用、IDE 或手机免费使用 Claude Code、Codex、Pi 和 OpenCode,支持语音输入。
Rust · ★ 36,726 · 🍴 3,670 · 📈 106 stars today
Your Personal AI super intelligence. A brain that builds a local-first memory of your life, a fantastic orchestrator of agent fleets and workflows, and a deep researcher.
中文介绍 个人 AI 超智能,构建本地优先的生命记忆,协调代理群和流程,进行深入研究。
JavaScript · ★ 242,544 · 🍴 36,718 · 📈 427 stars today
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
中文介绍 Claude Code、Codex、Opencode、Cursor 等的代理性能优化系统,提供技能、本能、记忆、安全和研究优先的开发。
TypeScript · ★ 69,060 · 🍴 8,273 · 📈 134 stars today
🌊 The original agent meta-harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory, self-learning intelligence, RAG integration, and native Claude Code / Codex / Hermes and many more Integrated
中文介绍 智能多玩家蜂群部署、自主工作流协调和对话式 AI 系统构建的原生元代理工具,具有自适应记忆和自学习智能。
★ 31,273 · 🍴 3,363 · 📈 223 stars today
A curated collection of 1000+ agent skills from official dev teams and the community, compatible with Claude Code, Codex, Gemini CLI, Cursor, and more.
中文介绍 1000+ 个代理技能集合,兼容 Claude Code、Codex、Gemini CLI、Cursor 等,适用于官方团队和社区。
Python · ★ 24,633 · 🍴 2,577 · 📈 423 stars today
Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.
中文介绍 将技术书籍 PDF 转换为 Claude Code 技能,便于学习和使用。
Rust · ★ 65,952 · 🍴 3,126 · 📈 95 stars today
Unofficial Bitwarden compatible server written in Rust, formerly known as bitwarden_rs
中文介绍 非官方的 Bitwarden 兼容服务器,使用 Rust 编写,提供密码管理服务。
Python · ★ 923 · 🍴 126 · 📈 257 stars today
Community plugin marketplace for Claude Cowork and Claude Code. Read-only mirror — submit plugins at clau.de/plugin-directory-submission.
中文介绍 Claude Cowork 和 Claude Code 的社区插件市场,提供插件提交和阅读功能。
HTML · ★ 134,405 · 🍴 14,062 · 📈 593 stars today
A list of SaaS, PaaS and IaaS offerings that have free tiers of interest to devops and infradev
中文介绍 列出对开发人员有价值的免费 SaaS、PaaS 和 IaaS 服务。
Python · ★ 129,367 · 🍴 15,254 · 📈 179 stars today
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
中文介绍 功能强大且模块化的扩散模型 GUI、API 和后端,具有图形/节点界面。
Python · ★ 234,963 · 🍴 47,334 · 📈 519 stars today
The agent that grows with you
中文介绍 随用户成长而发展的智能代理。
👍 16
Long-context modeling is a pivotal capability for Large Language Models, yet the quadratic complexity of attention remains a critical bottleneck, particularly during the compute-intensive prefilling phase. Our previous work, FlashPrefill, mitigates this cost through instantaneous pattern discovery a
👍 2
Large language model agents can adapt to complex tasks by constructing workflows at inference time, but procedures discovered in one episode are usually discarded after execution. Existing skill libraries provide reusable executable routines, but are typically assembled offline and do not grow from
👍 6
Multifingered grasping is a crucial robotic skill, but current deep-learning grasp planners often struggle to generalize to new objects because they are trained on limited, object-specific datasets. We introduce a fundamentally different approach, grounded in the observation that the gripper and the
👍 61
Software increasingly functions as part of the scientific instrument itself, making failures in scientific code capable of compromising not only program behavior but also the evidence underlying scientific conclusions. Yet existing evaluations of coding agents largely emphasize aggregate task succes
👍 17
Large language model agents have made substantial progress in code generation, yet most existing systems assume a predefined repository architecture. This assumption does not hold in zero-to-all code generation, where an agent must construct an entire software project directly from natural-language
👍 10
Large language models often fail to answer questions about a bounded document collection when the source documents are not retrieved at inference time. We study this setting as document knowledge internalization: converting a fixed corpus into usable parametric knowledge for retrieval-free question
👍 31
Memory has become a key component of large language models, enabling them to retain information and learn from long-term interactions. However, existing memory benchmarks mainly evaluate whether information is correctly extracted, stored, and retrieved, while largely overlooking how retrieved memori
👍 254
LLM agents learn by interacting with environments, yet these environments are hand-built and static: blind to an agent's weaknesses, and quickly left behind as it improves. While recent environment generation methods attempt to address this, they require domain-specific pipelines, rely on expensive
👍 8
Customer-service LLM agents must follow organizational policy when acting on a user's behalf. Compliance failures arise from either forbidden actions, such as granting an ineligible change, or omitted procedural requirements, such as identification or confirmation. Runtime safeguards can intervene o
👍 114
Training terminal agents requires scalable executable supervision, yet synthesizing high-quality terminal tasks remains challenging. Each task couples an instruction, an initialized environment, a reference solution, and an executable verifier; if these artifacts are generated from inconsistent assu
👍 11
Using reinforcement learning to post-train joint video-audio generation models requires a reward signal. Existing methods construct this reward by combining metrics for individual quality dimensions, including audio quality, visual fidelity, and synchronization. However, these metrics evaluate perce
👍 2
Object detectors often produce over-confident predictions for objects outside their training categories, leading to so-called out-of-distribution (OoD) hallucinations. Existing approaches for detecting or mitigating such hallucinations typically either construct scoring functions directly over learn
👍 6
Agent frameworks increasingly package procedural knowledge as skills: instruction files an agent reads on demand, while public libraries now hold thousands of them. Which skill to read has thus become a decision the policy itself makes in the middle of an episode, yet no existing signal trains it. W
👍 92
Reinforcement learning (RL) has emerged as a powerful approach for improving reasoning in language and vision-language models, yet its strongest successes still depend heavily on ground-truth supervision (e.g., verifiable reward). Such annotations are costly to obtain and become increasingly scarce
👍 15
JEPA-style latent world models can use Euclidean distance to a goal latent as the cost for model-predictive control (MPC). Strong decoding of task variables, however, does not guarantee that this particular cost ranks candidate action sequences by real task progress. We call the latter property deci
👍 13
Take three frontier mixture-of-experts models (Alibaba, OpenAI, NVIDIA; 3.6-4.0B active parameters each) and fine-tune them to reason in a low-resource language. On accuracy benchmarks almost nothing happens, and the benchmark itself is noise at this scale: changing only the random seed moves the sc
👍 8
Humans continuously learn from experience, whereas conventional large language model (LLM) evaluations ignore the models' ability to improve through inference-time interaction. In this paper, we study how LLMs learn from iterative experience at test time, a setting we refer to as Chain-of-Experience
👍 3
LiDAR scene completion is a key component of 3D perception in autonomous driving, where the scene must be completed in real time to be usable in downstream tasks. Existing approaches typically follow an initialize-and-refine paradigm, in which a coarse initialization of the scene is first constructe
👍 142
Embodied agents are increasingly used to close the gap left by end-to-end policy models. Yet the agentic path has not realized closed-loop learning in physical execution: existing harnesses remain largely open-loop, following fixed skills during rollout and reflecting only after an episode completes
👍 4
Language-specific competency (LSC) is the phenomenon of a language model performing better or worse depending on the language of the prompt. In other words, a language model outputs different (and potentially incorrect) responses to the same semantic query when prompted in different languages. Prior
👍 4
LLM-based agents can act on behalf of a user to access cloud services, call tools, or invoke agents. At session start, the agent's permissions are set but remain static, and each request is evaluated independently, without considering prior actions. Within its permissions, an agent may act contrary
👍 17
Popular facts are memorised more deeply during pretraining and resist removal longer than rare ones, yet existing LLM unlearning methods apply uniform gradient pressure regardless of training-data frequency. We propose the AdaPop (Adaptive Popularity) method, which combines local token confidence wi
👍 14
High-quality creative writing data for large language models (LLMs) remains dominated by story-centric data, limiting models' ability to follow the structural and functional conventions of diverse creative formats. We propose an attribute-guided genre expansion framework for scaling creative writing
👍 11
Should you replace your text-embedding pipeline with a large language model? We answer this with a controlled, cost-aware comparison of ten LLMs across six families and 26 embedding models (118M to 14B parameters) on 37 tasks spanning classification, semantic textual similarity (STS), clustering, pa
👍 7
LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot distinguish command-generation errors from failures introduced after generation. QuoteBench measures this boundary with exact final-state validation on 5
👍 8
Long-form video understanding encompasses tasks that go beyond retrieving isolated events, including tracking an evolving narrative and interpreting social meaning that may remain implicit. However, existing benchmarks rarely evaluate these capabilities jointly, particularly in high-context, non-Eng
👍 30
Agent Skills are today either hand-authored or produced in a single LLM generation pass, and consequently possess no closed loop through which they might improve from the interaction failures they actually cause. Recent work does close this loop, but derives its feedback from single-turn question-an
👍 88
Recent advances in foundation models have enabled AI scientists to automate increasingly complete research workflows, from hypothesis generation and code execution to manuscript preparation. Yet workflow coverage alone does not provide access to the full evidence on which scientific discovery depend
👍 11
Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the executable scaffold surrounding the model---the harness---is typically treated as a fixed artifact after deployment. This work studies an alternative where the harness is task-specific and continuously
👍 3
Token-level hallucination detectors score each token independently from a single signal, and fail exactly when the generating model is confidently wrong. This paper instead treats hallucination as a temporally extended span and detects it by sequence labeling: each token is scored from a 33-dimensio
@AndrewYNg · 1.8M 粉丝 · 62.4K 阅 · 1.2K 赞 · 189 转
I previously wrote about our AI Engineering Skills Map, with the highest level skills being (i) Building and deploying AI applications, (ii) Software engineering fundamentals, (iii) Using coding
中文介绍 AndrewYNg分享AI工程技能图谱,聚焦于构建和部署AI应用,强调高级技能如软件工程基础和代码使用。
Did you think RSI stopped at model training?
中文介绍 模拟技术发展迅速,成本降低100倍,速度提升1万倍,成为热门领域。
Models keep absorbing the harness into their weights — soon, it will be a harness for human attention rather than for the model.
中文介绍 模型吸收人类注意力机制,未来将更关注人类注意力而非模型自身。
Simile’s CEO about his journey from the viral Generative Agents to creating 8 Billion Digital Twins of every living human... and why it’s gone from fun exploration to very serious business.
中文介绍 Simile AI CEO分享其从生成代理到创建80亿数字双胞胎的历程,业务从探索转为严肃。
Google DeepMind partners with game studios to prototype breakthrough AI gameplay.
中文介绍 DeepMind与游戏工作室合作,原型突破性AI游戏。
Yes, we’re confused too.
中文介绍 Poolside获得12亿美元投资,创始人获得10亿美元,员工获得60亿美元,Infraco扩展至70吉瓦新云服务。
**Ox Alpha** emerged as a mystery model with strong coding and agentic performance, likely a **Zhipu/GLM-family** model such as **GLM-5.3 Vision**. Analysts suggest its gains come from post-training and infrastructure improvements rather than sheer size, based on the **743B base** of **GLM-5.2** wit
中文介绍 Ox Alpha成为神秘模型,表现优异,分析师认为其增长来自训练和基础设施改进。
中文介绍 语音识别基准优化测量。
中文介绍 ChatGPT Apple Messages、Anthropic会议记录器、Mistral Agentic Search等AI产品发布。
Matt Pocock tells us about his /wayfinder skill, for greenfield projects or for when the way forward is unclear.
中文介绍 Matt Pocock介绍他的/wayfinder技能,用于规划模糊不清的项目。
中文介绍 LFM2.5-DSpark推理速度提升至3.2倍。
“Runaway” AI, “rogue” agents, and “autonomous” actors—the current rhetoric would have you believe that AI agents are not only awake and aware, but angry at their creators. Prominent tech leaders such as Demis Hassabis, Dario Amodei, and Sam Altman push for regulation of these seemingly “superhuman”
中文介绍 AI意识辩论陷入误区,技术领袖呼吁监管。
Each day, an airline transports tens of thousands of passengers on hundreds of flights. Often these are not straightforward point-to-point routes, with passengers requiring multiple connections. The airline can consider potentially hundreds of variables to price each of these journeys: demand, seaso
中文介绍 市场模型解锁隐藏的收益流。
Introducing AI Futures, a new OpenAI blog exploring how transformative AI could reshape power, governance, the economy, and individual freedom.
中文介绍 OpenAI推出AI未来博客,探讨AI如何重塑权力、治理、经济和个人自由。
**OpenAI** and **Anthropic** expanded their agent platforms with new desktop features, collaborative editing, and composable APIs like Skills and Files API. **OpenAI** rolled out memory and workflow features in the EEA, UK, and Switzerland. **AT&T** revealed that 40% of employee AI usage routes to o
中文介绍 OpenAI和Anthropic扩展代理平台,推出新桌面功能、协作编辑和可组合API。AT&T透露40%员工...
Every lab CEO is on X now
中文介绍 Z.ai CEO Jie Tang讨论GLM 5.3和新的后训练扩展定律。
Leah Stewart lost her arm in a great white shark attack at Sydney’s Coogee beach. Follow today’s news live Get our breaking news email, free app or daily news podcast Joyce steps in on One Nation’s migration target Barnaby Joyce defended a fellow One Nation MP over comments suggesting the minor part
中文摘要 悉尼库吉海滩发生大白鲨袭击,Leah Stewart失去手臂。
PM will use independence day trip to announce sharing of classified information about Storm Shadow components Britain will boost Ukraine’s capacity to make its own long-range missiles, Andy Burnham will announce on Monday, as he travels to Kyiv for the country’s independence day on his first foreign
中文摘要 英国首相宣布将向基辅提供风暴阴影导弹技术信息,以增强其制造远程导弹的能力。
Dump site in Conakry, which minister had just promised to move, collapsed after heavy rain in west African state A landslide at a huge waste dump in Guinea’s capital has killed 30 people, the government said on Sunday, after heavy rains overnight prompted it to collapse, engulfing nearby tents and
中文摘要 几内亚首都科纳克里一垃圾山发生滑坡,造成30人死亡。
Barcelona begin their La Liga title defence and bid for a third successive crown with a 5-0 victory at Elche.
中文摘要 巴塞罗那5-0大胜埃尔切,开启西甲冠军卫冕之路。
The US president says he will impose 'crushing measures' on Tehran.
中文摘要 美国总统特朗普表示将对伊朗实施“压倒性措施”。
Serbia requested assistance from the European Union and Russia to help bring intense wildfires under control.
中文摘要 塞尔维亚请求欧盟和俄罗斯协助控制严重的森林火灾。
Ambassador to Turkiye says US policy of recognising Israel's claimed sovereignty over Syrian territory is 'unchanged'.
中文摘要 美国驻土耳其大使Tom Barrack撤回关于叙利亚戈兰高地被占领的评论。
In a recent article for Foregin Affairs, Mark Leonard makes that argument that America's leading role in the world order is fading. He discusses his argument with NPR's Danielle Kurtzleben.
中文摘要 马克·伦纳德在《外交事务》杂志中提出,美国在世界秩序中的领导地位正在衰落。
A mound of waste collapsed at a landfill in Conakry, Guinea, after heavy overnight rains, engulfing nearby homes.
中文摘要 几内亚首都科纳克里一垃圾山在暴雨后坍塌,造成30人死亡。
War, climate shocks and trade disputes are pounding the world’s breadbaskets and pushing the global food system into dangerous territory.
中文摘要 战争、气候冲击和贸易争端正在打击全球粮食产区,将全球食品体系推向危险境地。
Israel is moving forward with its ‘E1’ plan. What is it and why could it threaten the future of a Palestinian state?
中文摘要 以色列推进“E1”计划,可能威胁巴勒斯坦国的未来。
The president took a ceremonial lap of the track before the race, which is the culmination of a summer of events to mark the country's 250th birthday.
中文摘要 特朗普为华盛顿街道上的IndyCar比赛挥舞绿色旗帜,庆祝国家250岁生日。
Winds reaching around 120 km/h swept through Forlì in Italy’s Emilia-Romagna region, overturning four light aircraft.
中文摘要 意大利艾米利亚-罗马涅地区Forlì的强风将四架轻型飞机吹翻。
Race cars take over central Washington, DC, as the Trump-created Freedom 250 Grand Prix debuts.
中文摘要 赛车在华盛顿地标前高速行驶,作为特朗普创立的Freedom 250大奖赛的一部分。
Gaza’s hospitals face oxygen shortages, putting the lives of premature babies and critically ill patients at risk.
中文摘要 加沙地带医院面临氧气短缺,早产儿和危重患者的生命受到威胁。
The review of how rates are calculated in England and Wales could lead to reform of the system.
中文摘要 英格兰和威尔士的酒吧和酒店税率计算方法正在进行审查,可能导致税率制度改革。
A survey on attitudes towards accents suggests that so-called "accent bias" is still an issue in many workplaces.
中文摘要 一项关于口音态度的调查显示,所谓的口音偏见在许多工作场所仍然是一个问题。
Roehampton University offers students £2,000 a year to play esports alongside their studies.
中文摘要 罗汉普顿大学为学生在学习的同时玩电子竞技游戏提供每年2000英镑的报酬。
Fifth consecutive full-year loss exposes tension in mandate of lender set up by SNP-led government
中文摘要 苏格兰国家银行连续第五年录得1380万英镑亏损,暴露了由SNP领导政府设立的贷款机构职能上的紧张关系。
Andy Burnham is traveling to Ukraine on Monday for his first international trip as British prime minister, in an attempt to show he’ll maintain diplomatic and military support for Kyiv as Russia’s war enters a decisive phase.
中文摘要 英国首相Burnham首次访问乌克兰,试图展示英国将继续对基辅提供外交和军事支持。
Hong Kong’s ambitious plan to create a technology hub near the mainland Chinese border faces an early test after the first residential sales in its Northern Metropolis got off to a middling start, highlighting the hurdles to turning the area into a new economic engine.
中文摘要 香港在边境附近建立科技园区的雄心计划面临初步挑战,北部都会区首次住宅销售表现平平,凸显将地区转变为新经济引擎的障碍。
The Taiwan dollar’s strongest monthly rise in over a year is running into a wall of headwinds, sowing doubts among analysts on the sustainability of its rebound.
中文摘要 分析师表示,台币在8月份强劲反弹后,正面临重重困难,引发对其反弹可持续性的怀疑。
Regal Partners said Phil King will start a transition to retirement, marking the exit of one of Australia’s most high-profile hedge fund managers.
中文摘要 Regal Partners表示,基金经理Phil King将开始退休过渡,标志着澳大利亚最著名对冲基金经理之一的离职。
Shein Global Holdings Ltd. is seeking to raise as much as HK$13.9 billion ($1.8 billion) in its Hong Kong initial public offering, as it enters the final stretch of an arduous journey to go public.
中文摘要 Shein Global Holdings Ltd. 正在寻求通过其香港首次公开募股筹集高达139亿港元(18亿美元),以完成其艰难的上市之旅。
Oil edged lower in at the start of a key week for markets, with traders focused on the US plan to economically isolate Iran and the Federal Reserve’s annual gathering.
中文摘要 油价小幅下跌,阿里巴巴的资金计划让AI成为市场关注焦点。
Also in this newsletter: Alibaba announces $10.2bn share placement and Anthropic’s best AI model struggles to attract users
中文摘要 凯文·沃什寻求安抚投资者情绪,同时经济压力迹象日益增多。此外,阿里巴巴宣布发行102亿美元股份,Anthropic的顶级AI模型难以吸引用户。
Rezoning has changed the face of Auckland though critics warn lessons hard to replicate
中文摘要 新西兰最大城市通过重新规划改变了面貌,但批评者警告说,这些经验难以复制。
4 回复 · 程序员 节点
4 回复 · Apple 节点
67 回复 · 程序员 节点
34 回复 · 程序员 节点
30 回复 · Apple 节点
26 回复 · Apple 节点
51 回复 · Apple 节点
6 回复 · Apple 节点
21 回复 · Linux 节点
6 回复 · Linux 节点
该源今日无内容。