DietrichGebert/ponytail
JavaScript · ★ 153,370 · 🍴 8,222 · 📈 1,289 stars today
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
中文介绍 通过模拟懒散的开发者思维,帮助AI在编写代码时避免不必要的复杂度,提升代码质量。
JavaScript · ★ 153,370 · 🍴 8,222 · 📈 1,289 stars today
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
中文介绍 通过模拟懒散的开发者思维,帮助AI在编写代码时避免不必要的复杂度,提升代码质量。
JavaScript · ★ 75,283 · 🍴 4,516 · 📈 705 stars today
The design language that makes your AI harness better at design.
中文介绍 提供AI在设计中更优的辅助,提升设计语言的智能化水平。
JavaScript · ★ 272,223 · 🍴 40,648 · 📈 954 stars today
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
中文介绍 为Claude Code等AI工具提供性能优化系统,通过技能、直觉、记忆、安全性和优先开发研究,提升工作效率。
TypeScript · ★ 16,808 · 🍴 813 · 📈 302 stars today
Build production-ready applications in TypeScript
中文介绍 使用TypeScript构建生产级应用程序,提供高效的开发环境。
Go · ★ 109,522 · 🍴 6,333 · 📈 505 stars today
🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.
中文介绍 通过减少AI代理使用的token数量,提升处理效率,减少65%的token使用。
Python · ★ 89,775 · 🍴 7,898 · 📈 1,683 stars today
Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, XiaoHongShu — one CLI, zero API fees.
中文介绍 为AI代理提供互联网全视图,支持Twitter、Reddit、YouTube等多平台内容读取和搜索,无需API费用。
TypeScript · ★ 24,675 · 🍴 6,439 · 📈 251 stars today
中文介绍 提供代码生成工具,但具体功能未描述。
TypeScript · ★ 95,562 · 🍴 8,459 · 📈 218 stars today
Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More
中文介绍 为每个代理提供跨会话的持久化上下文,通过AI压缩和回注相关上下文,提升会话效率。
TypeScript · ★ 10,558 · 🍴 1,282 · 📈 84 stars today
Agent workspace built on Cloudflare Workers for creating documents, building apps, and running agents with your company’s context and systems.
中文介绍 基于Cloudflare Workers构建的代理工作空间,支持文档创建、应用开发和以公司上下文运行代理。
JavaScript · ★ 100,828 · 🍴 10,594 · 📈 305 stars today
Production-grade engineering skills for AI coding agents.
中文介绍 为AI编码代理提供生产级工程技能,提升编码效率。
Shell · ★ 294,907 · 🍴 26,354 · 📈 578 stars today
An agentic skills framework & software development methodology that works.
中文介绍 提供代理技能框架和软件开发方法论,提升工作效率。
Shell · ★ 275,347 · 🍴 23,104 · 📈 750 stars today
Skills for Real Engineers. Straight from my .agents directory.
中文介绍 为真实工程师提供技能,源自开发者个人代理目录。
TypeScript · ★ 25,244 · 🍴 1,807 · 📈 256 stars today
Context window optimization for AI coding agents. Sandboxes tool output (98% reduction), persists session memory, and enforces routing across 17 platforms via MCP + hooks.
中文介绍 优化AI编码代理的上下文窗口,通过沙盒工具输出、持久化会话记忆和多平台路由提升效率。
TypeScript · ★ 112,150 · 🍴 14,228 · 📈 408 stars today
AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
中文介绍 提供AI代理工具包,包括统一的LLM API、代理循环、TUI和编码代理CLI。
Python · ★ 45,201 · 🍴 4,898 · 📈 211 stars today
Developer-first error tracking and performance monitoring
中文介绍 提供开发者优先的错误追踪和性能监控服务。
TypeScript · ★ 149,216 · 🍴 25,359 · 📈 127 stars today
Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.
中文介绍 提供终端中的代理编码工具,理解代码库,通过执行常规任务、解释复杂代码和处理git工作流来提升编码速度。
Python · ★ 9,397 · 🍴 2,069 · 📈 192 stars today
中文介绍 提供生产级代理RAG课程,但具体内容未描述。
Python · ★ 8,729 · 🍴 1,523 · 📈 43 stars today
中文介绍 提供视频编辑工具,但具体功能未描述。
TypeScript · ★ 91,599 · 🍴 9,056 · 📈 234 stars today
The open-source CapCut alternative
中文介绍 提供开源的CapCut替代工具。
👍 21
Video generation models are increasingly being explored as world simulators for embodied planning and learning. To do so effectively, these models must not only generate visually appealing frames, but also predict how environments dynamically evolve when executing goal-directed actions. While evalua
👍 9
What makes great scientists great? Even as AI systems start to make progress on open problems, scientists remain far ahead of them at sensing which prior idea, buried in an ever-growing archive of research, a new problem needs. To study this skill, we draw on researchers who know firsthand which ear
👍 14
On-policy self-distillation has recently emerged as an effective approach for improving language-model reasoning by supervising students with a frozen or EMA version of themselves that receives privileged information. Its application to multimodal large language models (MLLMs), however, remains larg
👍 25
Looped Transformers achieve parameter efficiency by repeatedly executing a shared block across recurrent loops. Each loop yields an intermediate representation decodable for the same next token, yet standard decoding discards earlier states. Because earlier loops embody less computation, recurrence
👍 9
Generalization in large language models (LLMs) is the ability to produce consistent and semantically stable outputs when the same input is expressed in different ways. Existing work typically evaluates generalization through aggregate accuracy on a single prompt format, task, or set of variations, w
👍 5
Video Large Language Models (VideoLLMs) receive frames in sequential order and interpret how visual content evolves along the temporal axis, yet temporal reasoning remains a persistent weakness across architectures. Reversing the frame order of a video, a transformation that should invert temporal a
👍 34
Vision language model (VLM) agents can control robots through visual feedback and action primitives, but repeated model invocations and redundant observations incur substantial token overhead. We introduce PyRUA-Lean, an interactive code-execution framework that couples feedback-driven primitive com
👍 160
Streaming video LLMs must retain evidence before its relevance to future tasks is known and respond when sufficient evidence becomes available. The challenge is to form reusable factual memory without compromising real-time perception. We introduce OneStreamer, which jointly learns query-independent
👍 10
Recent video generation models can produce highly realistic videos from natural language instructions, with visual quality approaching cinematic standards. Existing evaluation benchmarks, however, predominantly assess visual quality, aesthetic appeal and physical plausibility, while paying limited a
👍 19
Extending a text embedding model to new modalities typically degrades text retrieval quality, and existing omni-modal embedders compensate with multi-billion parameters. We present Omni-Embed-Mini, a 0.9B-parameter model that maps text, speech, audio, images, video, and visually-rich documents into
👍 11
Retrieval-based speculative decoding (SD) drafts tokens by copying continuations from existing text, which suits coding agents that repeatedly reproduce code, logs, and earlier attempts. Yet existing methods fall short in agent pipelines: much of the reusable text is missing from their corpora or st
👍 8
LLMs are increasingly applied to cybersecurity workflows, where they are expected to translate analysts' intent into tool invocations. However, existing evaluations focus on knowledge-based assessments or end-to-end agentic tasks, and do not directly measure LLMs' ability to generate executable comm
👍 5
Video world models require persistent scene memory to maintain consistency during long-horizon video generation. Existing spatial memories accumulate RGB observations or latent features, increasing storage requirements as generation proceeds. We introduce Honeycomb, a video world model built on HexM
👍 7
LLM agents generate intermediate reasoning and actions token by token, making extended interactions slow and computationally expensive. Jev-style models offer fast probabilistic predictions over finite fields, but require those fields to be specified in advance. This requirement limits autonomous ta
👍 26
Furnished floor plans support real-estate visualization, interior design, and architectural workflows, yet automatic furnishing remains challenged by limited real-world data and the need to satisfy interacting geometric and functional constraints. We ask whether professional furnishing knowledge can
👍 17
Vision-Language Models (VLMs) have shown strong multimodal reasoning capabilities, yet whether they truly capture the physical consistency underlying real-world dynamics remains unclear. Existing benchmark paradigms often suffer from fragmented evaluation, focusing on isolated cognitive stages while
👍 44
Real-world embodied tasks, from everyday activities to professional procedures, require agents to act under physical constraints while tracking evolving object and task states. Tool use sits at the heart of such tasks, as many everyday and professional activities are tool-mediated. Understanding the
👍 14
On-policy self-distillation (OPSD) trains mathematical reasoning models using a privileged teacher that sees a reference solution and supervises student-sampled prefixes. Standard OPSD uses one fixed parameter setting at every state, but nearby settings may offer additional supervision. We find that
👍 6
An assistant that serves the same user over a long horizon has to answer from what that user has revealed: which preferences still hold, which were revised, and which constraints apply now. Retaining that information is not the same as acting on it, and the two are usually optimized as if they were.
👍 10
Coding agents are beginning to move beyond purely digital tasks to tackle physical-world challenges, particularly in robotics. Existing robotics benchmarks, however, primarily focus on the performance of individual artifacts, such as policies or controllers, offering limited coverage of coding agent
👍 11
Backdoor attacks can be implanted in Large Language Models (LLMs) during training, causing unwanted behaviour when a trigger appears in the input. Existing backdoor defences for LLMs attempt to remove the backdoor but inadvertently shift the model's output distribution to benign prompts, which can r
👍 23
Group Relative Policy Optimization (GRPO) is widely used to train reasoning language models, where it computes advantages by centering and normalizing rewards across rollouts of the same prompt. For multiple rewards, GRPO sums the reward components and normalizes the total reward by its within-group
👍 57
Masked diffusion models (MDMs) generate sequences by progressively unmasking several tokens per denoising step, but their reverse process is typically factorized over positions, limiting sample quality in the few-step regime where diffusion's speed advantage over autoregressive decoding matters most
👍 9
Audio-visual large language models (AVLLMs) have made remarkable progress in multimodal understanding and reasoning through interactions among visual, auditory, and linguistic information. However, recent studies show that AVLLMs face a critical challenge: source-confused grounding hallucination, wh
👍 153
On-policy learning has been argued to reduce catastrophic forgetting, produce sparser parameter updates, and improve generalisation. However, existing comparisons between supervised fine-tuning and reinforcement learning vary many factors simultaneously, making the contribution of rollout policy dif
👍 6
Before invoking external tools, an agentic LLM must select among a K-way action space: executing a call, seeking clarification, answering directly, or declining. While internal activation steering can alter these pre-execution decisions, conventional aggregate metrics obscure where altered states la
👍 5
Natural-language service requests can require a language-model decision before execution starts, consuming part of the request's latency budget. We integrate Jev's decision-oriented application programming interface (API) into edge service orchestration to reduce this overhead while retaining servic
👍 37
Multi-step agents are trained on flat action streams: SFT and RLVR weight every token uniformly and ignore the sub-procedures that recur across tasks, the hierarchy that lets humans plan top-down from reusable routines. This structure sits unused, and flat training uses each scarce trajectory less f
👍 11
Multidisciplinary tumor boards integrate multimodal clinical observations and longitudinal patient histories through specialist discussions, yet benchmarks rarely capture these real-world trajectories. We introduce OpenTumorBoard, a benchmark with 611 patient cases and 19,157 discussion turns across
👍 41
Decentralized multi-agent path finding (MAPF) with communication requires agents to reach individual goals without collisions under partial observability. Learnable policies trained on expert data provide an effective approach to this problem. However, when several coordinated joint actions are vali
@jakeserval · 1.1K 粉丝 · 406.0K 阅 · 505 赞 · 1.1K 转
There is a lot of talk about the jobs AI takes away and very little about the ones it creates. While pessimism might be easier (and certainly gets more clicks), I’m more optimistic. From where I sit,
中文介绍 探讨AI创造的新工作机会,而非取代的工作。
@0xROAS · 54.6K 粉丝 · 118.6K 阅 · 502 赞 · 56 转
Seedance 2.5 is one of the craziest "nano banana moments" of AI VIDEO in 2026. It's the model behind every viral video that you are seeing right now... and it's pretty simple to generate those video,
中文介绍 Seedance 2.5生成病毒视频教程,简单易用。
@0xSigil · 56.0K 粉丝 · 79.1K 阅 · 509 赞 · 72 转
6 months ago, the #1 AI was Claude Opus 4.6. Our model, Underdog 27B beats it while running on your Mac or iPhone. Most AI will run on your device, not in data centers. Intelligence that recently
中文介绍 介绍Underdog 27B AI模型,在个人设备上超越Claude Opus 4.6。
@UberEng · 66.9K 粉丝 · 54.0K 阅 · 529 赞 · 59 转
Introduction Uber’s rapid adoption of AI agents has fundamentally changed how teams interact with code, data, and operational systems. Early ad-hoc integrations with the MCP (Model Context Protocol)
中文介绍 Uber的MCP Gateway设计,AI代理如何改变团队与代码、数据交互。
@VibeMarketer_ · 37.8K 粉丝 · 31.4K 阅 · 504 赞 · 59 转
i'm going to show you how to build a content machine with Opus 5.5 that turns your ideas, research and expertise into articles, posts and newsletters in your own voice. from finding your next topic to
中文介绍 Opus 5.5构建内容机器教程,将想法转化为文章和通讯。
@goyal__pramod · 11.4K 粉丝 · 24.7K 阅 · 504 赞 · 59 转
Now I am aware you must have been looking forward to the second part of my CUDA blog (if you haven't read it yet, check it out!), but hey! A man can have varied interests. And I believe if you are
中文介绍 CUDA博客第二部分,介绍从头开始进行强化学习。
@ActionModelAI · 60.2K 粉丝 · 6.8K 阅 · 501 赞 · 439 转
OpenAI has chosen not to release GPT-6.1 Astra after the model failed to meet its safety standards. To be clear, GPT-6.1 Astra is not GPT-6.1 Sol, the model OpenAI released at DevDay. GPT-6.1 Sol has
中文介绍 OpenAI未发布GPT-6.1 Astra,暴露AI权限问题。
中文介绍 微软发布思考箱模型,数据库出现分歧。
a quiet day.
中文介绍 今日市场平静。
Learn how startups can choose GPT-6 models, tune reasoning effort, improve prompts and skills, coordinate tools, and prepare workflows for production.
中文介绍 OpenAI发布GPT-6系列模型构建指南。
Enterprise AI is no longer a future ambition. It is in full operational flight. Model capabilities are advancing faster than most organizations can absorb, while the cost of performance continues to fall. Globally, AI investment is set to reach $2.5 trillion in 2026, up 44% from the previous year. F
中文介绍 企业AI应用全面启动,预计2026年全球投资将达2.5万亿美元,增长44%。
中文介绍 Hugging Face开源AstaBrief快速报告生成模型。
After leading Meta’s Llama models, Ahmad Al-Dahle is now transforming Airbnb with AI — from how its teams develop products to how it serves guests.
中文介绍 Ahmad Al-Dahle领导Airbnb应用AI技术,从产品开发到客户服务全面升级。
On an afternoon in Seoul in March 2016, I watched a program I helped build put a stone on the fifth line of a Go board in what looked like a gift to its human opponent. Move 37 in game two of the five-game match looked so absurd that some commentators thought it was a…
中文介绍 LLM模型不具备推理能力。
the minimalist harness goes stable... and TypeScript!
中文介绍 Pi 1.0、Pi Durable和AIE NYC发布, minimalist harness稳定,支持TypeScript。
中文介绍 Hugging Face发布AutoSynthData,为企业代理生成训练数据。
We catch up with RLM first author Alex Zhang, MIT PhD, on Jev, PhD masxing, and the future of harnesses.
中文介绍 RLM第一作者Alex Zhang讨论Jev、PhD masxing和未来马具的发展。
Chatham Financial uses Codex and GPT-5.6 to build technology and redesign workflows, cutting trade validation from 30 minutes to under 4.
中文介绍 Chatham Financial利用OpenAI技术优化资本市场,将交易验证时间缩短至4分钟以下。
中文介绍 决策模型、Claude-shaped科学、OpenAI安全事件。
Advanced AI may matter most for the routine work behind breakthrough ideas. Explore why execution could shape the next economy and the pace of progress.
中文介绍 高级AI在突破性想法背后的日常工作中可能最为重要。
Albertsons Cos. is using ChatGPT Enterprise and the OpenAI API to help teams work faster and make grocery shopping easier for millions of customers.
中文介绍 Albertsons Cos.利用ChatGPT和OpenAI API优化零售业务。
... but you can’t try it yet unless you are “government users and trusted cyber defenders in the Fairwind Program”
中文介绍 Gemini 4 Argon发布,但仅限政府用户和受信任的网络安全防御者测试。
The US coastguard says it has dispatched air and surface crews to search for the plane near Nantucket.
中文摘要 波士顿至百慕大航班上失踪一架救护飞机,美国海岸警卫队已派员搜救。
Connie Chan, endorsed by Nancy Pelosi to succeed her in California, says she will vote to 'end a genocide in Gaza'.
中文摘要 佩洛西支持的美国候选人康妮·陈呼吁对以色列实施全面武器禁运。
Man charged with nine offences after vehicle struck Knights fans attending grand final farewell, leaving 10 injured including three children Get our breaking news email, free app or daily news podcast Ten people have been injured, one critically, after a car crashed into a crowd of people in Newcast
中文摘要 一辆汽车冲入新镇橄榄球赛粉丝人群,造成包括三名儿童在内的10人受伤,一人重伤。
President Trump says he's been having success winning races in Latin America. As Brazilian voters go to the polls, they're also deciding how much influence Trump should have in their country.
中文摘要 对于一些巴西人来说,这次选举是对特朗普影响力的投票。
High school students say their classrooms are overcrowded, teachers are missing and buildings need repairs. The protests have spread to cities across France, and some 5,000 people have been arrested.
Twenty-three years after President George W. Bush's campaign of "shock and awe," the U.S. military has completed its withdrawal from Iraq.
中文摘要 美国从伊拉克撤军行动已完成。
Brazil's election enters its final day of campaigning as Lula and Flávio Bolsonaro make their last pitches to voters. Brazilians explain what worries them, and why some still can't choose.
The aircraft lost communication with flight controllers after significantly dropping in altitude, according to flight data.
中文摘要 一架载有6人的医疗飞机在马萨诸塞州海岸失踪,飞机与地面失去联系后高度急剧下降。
Israeli settlers attack Palestinian farmers during olive harvest
中文摘要 以色列定居者在橄榄收获期间攻击巴勒斯坦农民。
The bridges, used by many commuters in the Ukrainian capital, are the latest target in a broad Russian bombing campaign against infrastructure.
中文摘要 俄罗斯对基辅的桥梁进行打击,威胁到重要的生命线。
The US president published a lawmaker's cell phone number as he called for the passage of a bill making DST permanent.
中文摘要 特朗普加大压力,要求美国共和党结束夏令时转换。
Funeral prayers were held at Gaza City’s Saint Porphyrius Greek Orthodox Church for a Palestinian mother and daughter.
中文摘要 加沙城圣波尔菲里乌斯希腊东正教教堂为在加沙空袭中遇难的母女举行葬礼。
The emperor is a symbol of Bulgarian national identity whose repatriation stoked tensions with Greece.
中文摘要 保加利亚沙皇萨穆埃尔遗骸在1000年后返回故土。
Youth-led movement demands resignation of chief election commissioner.
中文摘要 印度首席选举专员是否正在改变印度的政治格局?青年运动要求其辞职。
A photo went viral of a Palestinian boy hiding from Israeli forces on his way home from class in the occupied West Bank.
中文摘要 半岛电视台采访了一位在社交媒体上爆红的巴勒斯坦男孩,他在被占领的西岸上学途中躲避以色列军队。
President Luiz Inácio Lula da Silva and right-wing challenger Flávio Bolsonaro remain locked in a tight race as Brazil prepares to vote in Sunday’s election, according to a pair of new polls.
中文摘要 巴西总统大选进入拉锯战,卢拉与博尔索纳罗票数紧咬。
Borrowing is getting harder for companies.
中文摘要 企业借款难度加大,债务下降预示着借款人面临挑战。
The news doesn’t stop when markets close. Hosts David Gura, Christina Ruffini and Lisa Mateo bring clarity, context and a bit of humor to the weekend’s biggest headlines, LIVE from New York. Joined by The Atlantic Staff Writer Vivian Salama, “The Profiler” Online Job Scam Hunter Jay Jones, UMich Pro
中文摘要 纽约现场,Bloomberg This Weekend 带来周末重大新闻解析。
Flight 1073 was heading to Tel Aviv when the attack was launched in the cockpit
中文摘要 Flydubai 飞机副驾驶用斧头实施恐怖袭击。
Bloomberg Pursuits reporter Sarah Rappaport tells Bloomberg This Weekend that luxury travelers are increasingly seeking tiny hotels with as few as one to six rooms, combining the privacy of a villa with personalized high-end hotel service. Speaking with hosts David Gura and Christina Ruffini, Rappap
中文摘要 小型酒店成为奢华旅行的下一大趋势,奢华旅客追求私密与个性化服务。
Pointed offers a strategic twist to the news quiz format, testing not just players’ knowledge of the news but also their confidence in their answers. Join Bloomberg's Lisa Mateo, Christina Ruffini and David Gura as they play and check out the quiz for yourself at Bloomberg.com (Source: Bloomberg)
中文摘要 Bloomberg 周新闻测验带来战略转折,测试玩家知识和自信。
People starting to speak like AI chatbots, AI anxiety spilling into therapists' offices, NYC's AI office boom a bust for young job hunters and Tinder rolls out group dating for younger users. Join Lisa Mateo, David Gura and Christina Ruffini for a roundup of headlines you may have missed, but gotta
中文摘要 回顾你可能错过的新闻头条,包括人工智能焦虑和Tinder新功能。
Washington Institute research director Dana Stroul tells Bloomberg This Weekend that the US military buildup in the Middle East, including a third aircraft carrier, is expanding President Donald Trump’s options as diplomacy with Iran remains stalled and Washington weighs its approach to the Houthis
中文摘要 美国在中东的军事部署增加伊朗升级风险。
Fraud investigator Jay Jones, known online as “The Profiler,” tells Bloomberg This Weekend that AI is helping scammers create increasingly convincing fake recruiters and job listings that can lead victims to surrender money and sensitive personal information. Speaking with hosts David Gura and Chris
中文摘要 虚假职位使求职者面临身份盗窃风险,AI帮助骗子制造假招聘信息。
Atlantic Staff Writer Vivian Salama tells Bloomberg This Weekend that Egypt’s then-intelligence chief Abbas Kamel made an unusual 67-minute trip to Israel less than two weeks before the Oct. 7, 2023, attack to warn Israeli officials that intelligence indicated Hamas was mobilizing for a major assaul
中文摘要 埃及在10月7日前警告以色列哈马斯威胁。
Debate over artificial intelligence has largely focused on whether machines could become too powerful. Harvard philosopher Michael Sandel argues that a more immediate question is whether AI changes humans themselves: weakening authentic connection, independent thought and the ways people understand
中文摘要 哲学家担忧人工智能,AI可能改变人类自身,削弱真实联系和独立思考。
Chinese ecommerce company has been exploiting tax loophole that exempts small parcels from customs duty
中文摘要 中国电商公司Temu在英国的销售额超过1.71亿美元。
6 回复 · 程序员 节点
8 回复 · 程序员 节点
5 回复 · 程序员 节点
6 回复 · Apple 节点
16 回复 · 程序员 节点
11 回复 · 程序员 节点
9 回复 · 程序员 节点
29 回复 · Apple 节点
13 回复 · Apple 节点
17 回复 · Apple 节点
该源今日无内容。