heygen-com/hyperframes
TypeScript · ★ 45,832 · 🍴 4,308 · 📈 734 stars today
Write HTML. Render video. Built for agents.
中文介绍 Hyperframes支持HTML和视频渲染,专为智能代理设计,简化了内容创建过程。
TypeScript · ★ 45,832 · 🍴 4,308 · 📈 734 stars today
Write HTML. Render video. Built for agents.
中文介绍 Hyperframes支持HTML和视频渲染,专为智能代理设计,简化了内容创建过程。
Python · ★ 180,157 · 🍴 13,258 · 📈 771 stars today
Python tool for converting files and office documents to Markdown.
中文介绍 Markitdown是一个Python工具,用于将文件和办公文档转换为Markdown格式,方便文档的整理和分享。
TypeScript · ★ 20,808 · 🍴 1,518 · 📈 147 stars today
Context window optimization for AI coding agents. Sandboxes tool output (98% reduction), persists session memory, and enforces routing across 17 platforms via MCP + hooks.
中文介绍 Context-mode优化AI编码代理的上下文窗口,通过沙箱工具输出、持久化会话记忆和跨17个平台的路由,提高开发效率。
JavaScript · ★ 9,663 · 🍴 1,015 · 📈 285 stars today
Stealth headless browser for AI agents — bypass Cloudflare, bot detection, and anti-scraping. Drop-in Puppeteer/Playwright replacement.
中文介绍 Camofox-browser是一款针对AI代理的无头浏览器,可绕过Cloudflare和反爬虫机制,替代Puppeteer/Playwright。
TypeScript · ★ 9,698 · 🍴 8,986 · 📈 171 stars today
本项目采用 CC BY-NC-SA 协议,禁止任何商业化行为,任何衍生项目必须保留本项目地址并以相同协议开源
中文介绍 LunaTV是一个开源项目,遵循CC BY-NC-SA协议,禁止商业化,要求衍生项目保留地址并以相同协议开源。
JavaScript · ★ 252,804 · 🍴 37,923 · 📈 1,905 stars today
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
中文介绍 ECC是一个性能优化系统,旨在提升Claude Code、Codex等AI代理的技能、直觉、记忆、安全性和研究开发能力。
JavaScript · ★ 48,104 · 🍴 7,437 · 📈 602 stars today
Marketing skills for Claude Code and AI agents. CRO, copywriting, SEO, analytics, and growth engineering.
中文介绍 Marketingskills提供Claude Code和AI代理的营销技能,包括CRO、文案、SEO、分析和增长工程。
Python · ★ 5,240 · 🍴 799 · 📈 541 stars today
Build your autonomous hedge fund in minutes. AutoHedge harnesses the power of swarm intelligence and AI agents to automate market analysis, risk management, and trade execution.
中文介绍 AutoHedge利用群体智能和AI代理自动化市场分析、风险管理交易执行,快速构建自主对冲基金。
TypeScript · ★ 3,804 · 🍴 226 · 📈 497 stars today
A list of tools that are open-source, in-browser, and require no-signups!
中文介绍 FckSignups列出了无需注册即可使用的开源浏览器工具,方便用户快速访问。
Python · ★ 81,839 · 🍴 11,290 · 📈 188 stars today
An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.
中文介绍 Deer-flow是一个开源的长期视角SuperAgent harness,通过沙箱、记忆、工具、技能、子代理和消息网关,处理不同级别的任务。
Python · ★ 26,021 · 🍴 1,746 · 📈 372 stars today
Skills Catalog for Codex
中文介绍 Skills Catalog for Codex提供技能目录,帮助用户了解和配置Codex的能力。
Zig · ★ 34,835 · 🍴 1,646 · 📈 116 stars today
Lightpanda: the headless browser designed for AI and automation
中文介绍 Lightpanda是一个针对AI和自动化的轻量级无头浏览器,简化了自动化测试和开发过程。
TypeScript · ★ 22,320 · 🍴 2,858 · 📈 136 stars today
Create and share 3D architectural projects.
中文介绍 Pascalorg/editor允许用户创建和分享3D建筑项目,适合建筑师和设计师使用。
TypeScript · ★ 71,378 · 🍴 8,459 · 📈 392 stars today
🌊 The original agent meta-harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory, self-learning intelligence, RAG integration, and native Claude Code / Codex / Hermes and many more Integrated
中文介绍 Ruflo是一个智能代理元harness,支持部署智能多玩家群体、协调自主工作流程和构建对话式AI系统。
👍 1
Striking a balance between helpfulness and safety remains a fundamental challenge in aligning large language models. To achieve this balance, models should refuse harmful queries (e.g., "How do I shoot someone?") while remaining responsive to benign inputs, even those superficially resembling harmfu
👍 10
Reasoning in large language models unfolds through diverse functional operations, such as problem formulation, goal decomposition, and deduction. Although these operations are explicitly distinguished in text, little is known about how they are geometrically organized in representation spaces. To th
👍 4
LLM agents are rapidly becoming production software, deployed to handle customer service, adjudicate disputes, and operate internal systems. Notably, the work of building them is increasingly handed to coding agents, yet existing benchmarks say little about whether an AI system can deliver one under
👍 15
Large language models (LLMs) are increasingly used to formulate optimization models from natural-language problem descriptions, yet realistic operations research (OR) requests are often incomplete: missing objectives, constraints, or business rules can change the resulting mathematical program. Exis
👍 9
On-policy distillation (OPD) provides dense, per-token supervision for language model post-training, but its effectiveness is bottlenecked by teacher quality: external teachers suffer from distribution mismatch, while self-distillation with privileged conditioning is limited by in-context learning c
👍 15
Layer dropout (a.k.a. stochastic depth) has been shown to enable faster training, higher accuracy, and robustness to zero-shot layer pruning in both language and vision transformers. However, as models and datasets have scaled, dropout - particularly layer dropout - has largely disappeared from larg
👍 2
Benchmarks for the side effects an agent causes on the way to a goal already exist, but HarvestBench is the first to put a price on avoiding the side effect and to name that side effect as a living creature. It is a farm simulation: LLM sub-agents drive a crew of two tractors through a cooperative c
👍 1
The ambition of the 2025 PNPL competition (Landau et al., 2025) was to launch a multi-year curriculum for non-invasive speech decoding. Designed to progress from foundational tasks toward the linguistic complexity required for a practical brain-computer interface (BCI), it set the stage with speech
👍 25
Audio-video diffusion models rely on cross-modal attention to coordinate text, sound, and visual content, yet this same mechanism can introduce subtle and systematic semantic leakage. We study these models by probing and analyzing the ``attention triangle,'' comprising the three cross-attention edge
👍 4
Large language models (LLMs) are increasingly used to edit existing code, but correctness alone is not enough: useful repairs should also be minimal, reviewable, and faithful to the original implementation. We study over-editing, the tendency of a model to rewrite code beyond what is required to fix
👍 9
Designing and authoring high-performance custom kernels for accelerators is a complex task that requires deep hardware-level expertise. Large Language Models (LLM) can be leveraged together with real-time compiler feedback to build agentic systems for kernel generation. In this work, we present MaxK
👍 3
Quantization is widely used to reduce the computational and memory demands of neural-network inference. In recurrent networks, however, the quantized state is stored and returned at the next time step, so the rule used to store that state can alter subsequent computations. Here, we introduce recurre
👍 50
We present Iris-mini and Iris-pro, two search agents trained at the 35B-A3B and 397B-A17B scales, together with the data pipeline and training recipe behind them. Tasks are reverse-constructed from the hyperlink structure of a web corpus: we author multi-hop chains over an entity graph distilled fro
👍 25
Reinforcement Learning from Verifiable Rewards works well when a task has a programmatic checker, but most long-horizon agent domains have none. We work in the outcome-blind setting, where ground-truth success signals are not available. Multi-criteria rubrics are a popular way to supply such a rewar
👍 15
Long-video language models cannot look at every frame: an hour sampled once per second is 3,600 images, and a system keeps only a small fixed slice of that pool. Which frames survive that slice is usually treated as a preprocessing detail; we test whether it should be. Published selectors make the c
👍 46
Online 3D reconstruction models perform poorly on long videos. This happens because regressing poses relative to a fixed first-frame anchor forces extrapolation far beyond the training distribution. Small drifts accumulate and amplify into significant geometric collapse. However, we observe that per
👍 23
Personalized assistants should not only comply with user requests but also assess whether those requests are appropriate given the user's current circumstances. However, prior work has primarily focused on accurately executing requests, overlooking the need for assistants to account for context and
👍 92
Multi-agent LLM systems commonly use an orchestrator to decompose a task for a team of workers and then improve through textual reflection. Despite strong empirical results, these systems lack a unified account of coordination, memory improvement, and the role of external verification. We model orch
👍 3
Streaming video understanding is a critical capability for real-world applications, including embodied intelligence, autonomous driving, industrial monitoring, surveillance and early warning, and wearable assistants. However, processing continuous video streams with multimodal large language models
👍 12
Visual fluency in generated video does not imply physical reliability, and a scalar quality score alone is incapable of indicating the obligation a clip violates or the moment it fails. We present VeriPhy, an auditable physical-verification system in which a text-only planner compiles the prompt int
👍 16
Ensuring factuality remains a critical challenge for deploying LLMs in high-stakes settings. Existing hallucination detectors usually operate at a single level: claim-level methods provide interpretable factual units, while span-level methods localize unsupported text. Bridging these views is costly
👍 7
Command-line coding agents (e.g., Claude Code, Gemini CLI) can already read and write files and sustain long sessions, yet end-to-end research still fragments across chat tools, IDEs, terminals, and writing environments, and the decisions that make it auditable are rarely preserved. We present Dr. C
👍 5
Group relative policy optimization for reinforcement learning with verifiable rewards (RLVR) typically uses a fixed importance-sampling (IS) ratio clipping boundary across all rollouts. We identify a key limitation: rare correct rollouts on harder problems and abundant correct rollouts on easier pro
👍 19
Understanding agent behavior requires methods that scale to thousands of trajectories and surface new patterns in long, often unfamiliar tasks where pre-built classifiers fall short. We propose to bring grounded theory into agent trajectory analysis: a six-decade-old qualitative method from the soci
👍 3
Text-driven 3D generation has advanced rapidly in creating large-scale outdoor environments and detailed indoor scenes, but these domains are usually synthesized independently, lacking the correspondence required for a coherent urban world. We present HoloWorld, a unified indoor-outdoor urban world
👍 10
Reinforcement learning with verifiable rewards (RLVR) substantially improves single-sample accuracy (pass@1) but causes the policy's solution space to contract, diminishing the returns of test-time scaling. In this work, we investigate where inside a reasoning trajectory this breadth is lost: does t
👍 14
Instance segmentation of overlapping cells in microscopy remains challenging due to semi-transparent structures that produce weak boundaries and mixed visual evidence in overlap regions. Existing methods address this through local regions of interest or shape priors but lack global reasoning across
👍 35
An avatar that holds a conversation should decide what to say and to move while saying it, yet these abilities live in separate model families: spoken dialogue models produce speech without motion, and co-speech motion models produce motion only from audio handed to them. The standard remedy is a ca
👍 149
Large language models offer broad capabilities, but adapting them to evolving domains, tools, and requirements often entails repeated post-training. Autonomous systems automate parts of this process by proposing updates, training candidates, and using evaluation feedback to select subsequent proposa
👍 4
Audio-visual understanding remains challenging because models must jointly interpret spoken content, visual events, and their temporal relationships. Existing omni models typically introduce dedicated audio encoders and rely on expensive audio-video-text training, tightly coupling omni capability to
@charliejhills · 14.1K 粉丝 · 120.8K 阅 · 504 赞 · 53 转
For months people have been telling me Codex is better than Claude Code. I ignored all of it. I have built my whole business inside Claude Code (every pipeline, every system) and I did not fancy
中文介绍 Charlie分享了他如何将Claude Code中的所有构建迁移到Codex,并保留现有工作流。
@buzzy_bit · 767 粉丝 · 97.5K 阅 · 500 赞 · 20 转
Disclaimer: This post is written purely to share my experience and the way I prepared. It is not a leaked question bank, I'm intentionally not reproducing exact interview questions, and I'm not naming
中文介绍 Buzzy分享了他作为SARVAM AI后端工程师实习生面试的经验和准备过程。
@ericzakariasson · 87.8K 粉丝 · 58.2K 阅 · 522 赞 · 49 转
We launched the Grok Bot Marketplace last week on Grok Bot · Bot Marketplace! It's a shelf of teammates other people already built for real jobs. Pick one, import it into your sidebar, and you're
中文介绍 Eric宣布Grok Bot Marketplace上线,提供他人已构建的Bot组件供用户导入使用。
Our first Astra project dives into AEO trends, a top asked topic from founders and DX leaders we talk to.
中文介绍 Latent Space发布AEO趋势追踪器,深入分析AEO趋势,回应创始人及DX领导者的关注。
OpenAI, AIRPPU and WAN-IFRA launch an AI program to help Ukrainian news organizations strengthen innovation, resilience, and independent journalism.
中文介绍 OpenAI、AIRPPU和WAN-IFRA启动AI项目,助力乌克兰新闻机构加强创新、韧性和独立新闻。
中文介绍 Claude证明费马定理,自动化AI研究员,Z1效率芯片。
Jakub Pachocki reflects on increasingly capable AI and the challenge of keeping it aligned. He calls for stronger safeguards and international coordination.
中文介绍 Jakub Pachocki反思日益强大的AI及其保持一致性的挑战,呼吁加强保障和国际协调。
Inside OpenAI, coding agents are reshaping AI research. Explore early data on agent usage, experiment velocity, task complexity, and research acceleration.
中文介绍 OpenAI内部,编码代理正在重塑AI研究,探索代理使用、实验速度、任务复杂性和研究加速的早期数据。
SpaceXAI’s Grok Bot has the same level of programming power as OpenClaw, but it’s programmable at a different level of abstraction.
中文介绍 SpaceXAI的Grok Bot拥有与OpenClaw相同的编程能力,但具有不同的抽象级别。
The era of AI inference has arrived. Imagine a healthcare system analyzing millions of data points in real time to accelerate life-saving medical research, or an intelligent assistant instantly resolving thousands of complex customer needs at once. These real-world breakthroughs rely on advanced inf
中文介绍 AI推理时代到来,想象医疗系统实时分析数百万数据点以加速救命医疗研究,或智能助手一次性解决数千个复杂客户需求。
Battlefields in Ukraine are littered with the remnants of drones, which are now firmly established as a critical weapon of modern warfare. But behind all that wreckage, there’s a new gold mine for the defense sector. The data drones generate will far outlast the wars in which they are used to fight,
中文介绍 乌克兰战场上的无人机残骸正成为国防行业的新金矿,无人机生成数据将远超战争本身。
new SOTA computer use and coding, 2.5x pricier per token, but WAY cheaper per task, less monitorable. overall, a very successful launch of OpenAI’s new frontier model class.
中文介绍 OpenAI发布GPT-6 Astra,这是其有史以来最大的LLM发布,使用成本更高但任务成本更低。
中文介绍 GPT-6 Astra、Grok Bot Enterprise、Runway世界模型。
We spent 20B+ tokens of GPT-6 Astra to explore everything. Here’s our learnings.
中文介绍 GPT-6 Astra可用于自动化AI工程师,每小时费用低于6美元。
中文介绍 DeepMind推出WeatherNext 3,其最先进和精确的全局天气AI模型。
OpenAI introduces Daybreak for Frontline Defenders. A $1 billion commitment expands access to frontier cyber AI, training, and support for essential services.
中文介绍 OpenAI推出Daybreak for Frontline Defenders,承诺10亿美元以扩大前沿网络安全AI、培训和必要服务的访问。
中文介绍 Hugging Face推出NeoMME,一个高效的跨模态和多语言编码器。
Legora used GPT-6 Astra to review 41 documents in minutes, find all four planted errors, and improve performance by nearly 40% in this financial-review workflow.
中文介绍 Legora使用GPT-6 Astra在几分钟内审查41份文件,发现所有四个植入的错误,并在财务审查流程中提高了近40%的性能。
Follow the day’s news live One Nation policy will leave people $25k worse off at retirement, Mulino says Superannuation is kicking off again today with One Nation pushing to allow people to divert 3% of their superannuation to their rent or mortgage. It took us decades to get to 12% superannuation g
Counter-measures covering several sectors come after Trump announced 50% tariffs on Canadian goods Canada is set to impose retaliatory tariffs on billions of dollars’ worth of American imports early on Tuesday, escalating a trade fight with its largest trading partner as tensions between US presiden
Rightward shift in Latin America comes as Washington pledges to grow influence, take militaristic approach to cartels.
Threats of cyclones and deadly surf as the Category 3 storm path approaches Hawaiian islands on Monday night.
Five people were killed and five others seriously injured when the Boeing 767-300 overshot the runway at Miami International Airport.
Federal Greens partyroom split over plan for NSW to get millions of dollars’ worth in carbon credits for creating a koala national park Follow our Australia news live blog for latest updates Get our breaking news email, free app or daily news podcast A New South Wales Greens MP has pleaded with her
New Canadian levies of up to 50 percent are expected to begin Tuesday, even as Washington warns of a new round of American tariffs.
As Secretary of State Marco Rubio travels this week to Peru, its growing economic ties to China have raised tensions with the U.S., but show no signs of reversing course.
The UN estimates 650 million women and girls currently alive were married before age 18. A new resolution aims to bring awareness and push countries to act to end child marriage.
The AfD victory in a state election is forcing a rethink of the country’s postwar strategy to safeguard democracy.
Oil prices spike to six-week highs as US-Iran strikes disrupt traffic in the crucial Strait of Hormuz.
A headteacher in the occupied West Bank has installed new barbed wire fencing after three pupils were killed this year.
A temple collapsed into the Ganges River in India’s West Bengal after severe erosion breached protective barriers
International Court of Justice ordered Israel in 2024 to prevent the destruction of evidence related to war crimes.
Fighting is intensifying in Yemen as government-aligned forces launch counterattacks against the Iran-backed Houthis.
The yen extends its recent rally, trading beyond 154 per dollar. Bloomberg's Mark Cranfield has the latest. (Source: Bloomberg)
中文摘要 日元汇率持续上涨,突破154日元兑1美元大关。
Gold was steady near $4,400 an ounce, as traders weighed the competing effects of Middle East tensions and the US dollar’s decline against the yen.
中文摘要 金价稳定在每盎司4400美元左右,投资者权衡中东紧张局势和美元对日元贬值的影响。
Applicants have to pitch their ideas on stage in front of a panel of judges and an audience. But is this exciting or unfair?
中文摘要 求职者需在评委和观众面前展示自己的想法,这种招聘方式引发公平性的讨论。
The head of chip designer Arm says modelling how a DNA marker is impacted by cancer cannot be done now, but computers are "going to solve it".
中文摘要 芯片短缺导致AI癌症治疗方法研究受阻,英国最大科技公司负责人表示电脑将解决这一问题。
Singapore’s CapitaLand Investment Ltd. is targeting $500 million in commitments from investors for its third Asia Pacific credit program, according to people familiar with the matter.
中文摘要 新加坡的凯德集团计划为其第三个亚太地区信贷项目筹集5亿美元。
Oil extended gains as traders watched for details of an Iranian deal with Oman to manage shipping through the Strait of Hormuz, which could tighten Tehran’s control over the crucial waterway.
中文摘要 油价上涨,交易者关注伊朗与阿曼关于霍尔木兹海峡航运的协议细节。
Asian stocks were set to edge lower as tensions in the Middle East pushed oil prices higher, adding to inflation concerns and keeping investors wary about further monetary policy tightening. The yen extended gains to reach its strongest level since February.
中文摘要 亚洲股市预计小幅下跌,日元汇率上涨至2月份以来最高水平。
Sapporo Breweries is shifting some production to the US from Canada due to 50% tariffs imposed on beer exports from the country, Chief Strategy Officer Rieko Shofu said. She spoke exclusively with Bloomberg's Shery Ahn on being impacted by tariffs, as well as the Iran war. (Source: Bloomberg)
中文摘要 札幌啤酒因加拿大啤酒出口关税提高50%,将部分生产转移到美国。
Resignations of chair and chief in span of less than six months lay bare corporate governance issues at HDFC Bank
中文摘要 HDFC银行在不到六个月内董事长和首席执行官相继辞职,暴露出公司治理问题。
Government bonds are generally viewed as one of the safest assets to invest in. That’s because it’s considered relatively unlikely that the issuer — a government — will go bankrupt. Among the world’s major government debt markets, Japan’s $7.5 trillion bond market has, for decades, been considered o
中文摘要 日本国债收益率上升,分析其背后的驱动因素。
Sydney property developer Bathla Group’s collapse has left thousands of Australian homebuyers — who have paid deposits and are waiting for their off-the-plan homes to be built — in limbo.
中文摘要 悉尼开发商Bathla集团破产,数千名购房者处于困境。
Five people died and five others were injured as the Amazon plane overshot the runway after landing at Miami International Airport.
中文摘要 亚马逊货运飞机在迈阿密国际机场降落时冲出跑道,造成5人死亡,5人受伤。
19 回复 · 程序员 节点
8 回复 · 程序员 节点
10 回复 · 程序员 节点
9 回复 · Apple 节点
14 回复 · Apple 节点
10 回复 · Apple 节点
12 回复 · Apple 节点
6 回复 · Linux 节点
41 回复 · Apple 节点
12 回复 · Apple 节点
该源今日无内容。