TechCrunch · 07/23 04:49
美国财政部长警告,白宫指控中国月之暗面公司「提炼」Anthropic的Fable模型开发Kimi K3,美国政府可能因此对中国AI公司实施制裁,加剧科技领域的地缘政治紧张。
推荐理由:此事件揭示AI技术竞争中的国际地缘政治风险和模型版权争议,可能对全球AI产业发展产生深远影响。
TechCrunch · 07/23 04:49
美国财政部长警告,白宫指控中国月之暗面公司「提炼」Anthropic的Fable模型开发Kimi K3,美国政府可能因此对中国AI公司实施制裁,加剧科技领域的地缘政治紧张。
推荐理由:此事件揭示AI技术竞争中的国际地缘政治风险和模型版权争议,可能对全球AI产业发展产生深远影响。
The Decoder · 07/23 03:33
Anthropic因从盗版数据库下载作品而支付15亿美元和解金,创集体诉讼历史新高。此案焦点并非AI训练本身,而是数据来源的合法性,此前已有法官裁定合法获取作品用于AI训练属合理使用。
推荐理由:此和解案凸显AI模型训练数据来源的法律风险和版权合规的重要性,为AI行业敲响警钟。
GitHub Trending
OmniRoute 是一个免费开源AI网关,通过单一API端点集成268+提供商、500+模型(含50+免费),支持Claude、GPT、Gemini等,旨在简化多模型调用并大幅降低开发成本,还通过RTK+Caveman技术实现15-95%的数据压缩。
推荐理由:为开发者提供一站式AI模型调用解决方案,显著提升开发效率并降低成本,是AI应用部署的实用工具。
Wired AI · 07/23 03:01
面对Anthropic和OpenAI前沿模型访问受限,中国实验室积极推出开源AI替代方案,强调其稳定性、可访问性和日益增强的能力,对硅谷在AI领域的传统发展模式构成挑战。
推荐理由:反映全球AI竞争格局的变化,中国开源AI力量的崛起为全球开发者提供了更多选择,值得关注其长期影响。
Claude (YouTube) · 07/23 01:05
Claude在YouTube发布短视频,深入探讨人工智能系统在学习和训练过程中,如何逐步发展出独特的行为模式,即其「性格」。该视频以通俗方式解读AI个性化的复杂机制。
推荐理由:以轻松形式科普AI行为模式的形成,有助于公众更好地理解AI系统,引发对AI伦理和设计的思考。
Product Hunt · 07/23 04:50
Remote OpenClaw平台为AI编程代理提供强大的基础设施支持,集成超过13,000个MCP服务器及丰富的技能与插件,旨在显著增强AI代理的编码、开发和自动化能力。
推荐理由:为AI编程代理提供全面的基础设施和资源,是推动AI Agent开发与应用的关键工具。
Riley Brown (YouTube) · 07/22 06:27
YouTube博主Riley Brown分享了他将OpenAI Codex(代码生成AI)转化为业务增长引擎的实践经验。视频详细阐述了如何利用AI的代码生成能力,实现企业运营和开发的效率提升,以加速业务发展。
推荐理由:通过个人实战案例,展示AI工具在业务增长中的巨大潜力,为开发者和创业者提供了实用的应用指南。
V2EX · 07/22 19:28
V2EX社区讨论了如何安全地请美国用户代为申请和支付Claude AI服务,同时确保代理方银行卡/信用卡信息的安全。讨论旨在找到不涉及个人敏感金融信息泄露的可靠支付方案。
推荐理由:针对部分地区用户访问AI服务难题,提供了社区讨论与解决方案,对有类似需求的用户具有参考价值。
X 推文 (AttentionVC) · 07/22 02:23
Google AI Studio宣布,面向生产环境的Gemini 3.6 Flash和3.5 Flash-Lite模型已正式全面上市。其中,Gemini 3.6 Flash在复杂任务处理上表现出更强性能,为开发者提供了更强大的选择。
推荐理由:Google新一代Flash模型全面上市,为需要高性能和轻量级AI解决方案的开发者提供了更多选择,值得关注和应用。
HuggingFace Trending Papers · 07/22 01:51
该论文提出ISO,一个专为可验证奖励强化学习(RLVR)设计的优化堆栈。它旨在解决将奖励反馈转化为语言模型权重更新的优化难题,有望进一步提升大模型的推理能力和RLVR技术的应用。
推荐理由:为推动RLVR在语言模型推理中的应用提供了新的优化框架,对AI研究和技术突破具有重要意义。
Treasury Secretary Scott Bessent warned the U.S. government could sanction Chinese AI companies after White House officials accused Moonshot of distilling Anthropic's Fable model to develop Kimi K3.
中文介绍 美国财政部长斯科特·贝森特警告,在美国白宫官员指控中国AI公司月之暗面(Moonshot)通过“提炼”Anthropic的Fable模型来开发Kimi K3后,美国政府可能会对中国AI公司实施制裁。
Tesla's 26% boost in revenue wasn't enough to offset rising operating expenses and capital expenditures as it pushes to launch a new generation of products.
中文介绍 尽管特斯拉收入增长26%,但运营费用和资本支出飙升,未能抵消成本。同时,其新一代产品如Cybercab、Semi和Megapack的生产计划有所推迟。
A closely watched social media addiction lawsuit that had been set to go to trial next week has been dropped after the plaintiff voluntarily dismissed his claims against Meta, leaving none of the major tech companies facing trial in the case.
中文介绍 一起备受关注的针对Meta的社交媒体成瘾诉讼案已被撤销。原告自愿撤回对Meta的诉求,这意味着目前没有主要科技公司面临此类案件的审判。
A Tesla Model 3 electric car is seen during the China International Supply Chain Expo (CISCE) in Beijing on July 16, 2025. (Photo by Jade GAO / AFP) (Photo by JADE GAO/AFP via Getty Images) | AFP via Getty Images After a dismal two years of weakening demand, falling sales, and damage to its brand by
中文介绍 经历两年需求疲软和销量下滑后,特斯拉的营收开始回升,但公司的盈利能力仍然疲弱。
SoundCloud has acquired decentralized music platform Nina Protocol, months after the startup announced it would shut down. The deal brings Nina’s artists, editorial archive, and music discovery tools to SoundCloud as the company continues expanding its platform for independent musicians.
中文介绍 SoundCloud收购了去中心化音乐平台Nina Protocol,而Nina Protocol在数月前才宣布关闭。此次收购将把Nina的艺术家、编辑档案和音乐发现工具整合到SoundCloud,以继续拓展其面向独立音乐人的平台。
Anthropic has to pay $1.5 billion to book authors, the largest copyright settlement in class action history. But the payout is for downloading roughly 482,460 works from piracy databases, not for AI training itself. Judge Alsup had previously ruled that AI training on legally obtained books is "tran
中文介绍 Anthropic须向图书作者支付15亿美元,这是集体诉讼史上最大的版权和解金。但这笔款项是因从盗版数据库下载约482,460部作品而支付,并非用于AI训练本身。此前法官Alsup曾裁定,在合法获取的作品上进行AI训练属于合理使用。
The iPad Pro M5 is one of the iPads discounted for the sale. | Image: The Verge A number of Apple products got more expensive last month, so we’re happy to find deals wherever and whenever we can. If you’re searching for a high-end iPad, one of the more notable deals currently happening is on the 13
中文介绍 上个月价格上涨后,包括iPad Pro M5在内的一些苹果产品目前正在打折。这为寻求高端iPad的用户提供了一个暂时的价格优惠机会。
Code found in an iOS 27 beta would allow Apple to put a financed iPhone in "Restricted Mode" if it detects any missed payments, 9to5Mac reports. The finding follows a story from Bloomberg earlier this week claiming Apple is preparing to launch a new "Apple Upgrade" financing program for leasing new
中文介绍 据9to5Mac报道,iOS 27测试版代码显示,苹果可能会在用户未按时支付分期付款时,将分期购买的iPhone置于「限制模式」。此前彭博社曾报道,苹果正准备推出新的「Apple升级」分期计划。
OpenAI made a mistake setting up what it called a “highly isolated” testing environment and sandbox. According to cybersecurity experts, that human mistake is what made the AI-powered attack on Hugging Face possible.
中文介绍 网络安全专家表示,OpenAI在设置其“高度隔离”的测试环境和沙盒时出现了人为失误,正是这一错误导致了针对Hugging Face的AI驱动网络攻击得以发生。
As access to Anthropic’s and OpenAI’s frontier models becomes more restricted, Chinese labs are pitching their open-source alternatives as stable, accessible, and increasingly capable.
中文介绍 随着Anthropic和OpenAI前沿模型的访问受限,中国实验室正推出其开源替代方案,并宣称这些模型稳定、易于访问且能力日益增强,挑战硅谷的AI发展模式。
Uber is also investing in Travis Kalanick's company Atoms, which has made gauzy claims about using industrial AI to modernize the world.
中文介绍 Uber创始人特拉维斯·卡兰尼克(Travis Kalanick)的机器人公司Atoms在新一轮融资中筹集了17亿美元,由a16z领投。Uber也投资了该公司,Atoms声称利用工业AI实现世界现代化。
Union previously warned automaker that any robot deployment must be negotiated.
中文介绍 现代汽车声称,其仿人机器人计划并非与罢工工人谈判的一部分。此前,工会曾警告这家汽车制造商,任何机器人的部署都必须通过谈判进行。
"The thing that the space needs is a company making $100 million a year of revenue," Science Corp. CEO Max Hodak said.
中文介绍 Science Corporation开发的视力恢复芯片已获得欧盟批准。公司首席执行官马克斯·霍达克(Max Hodak)表示,该领域需要一家年收入达到1亿美元的公司。
Yope, a fast-growing social app focused on private groups of friends and family, has raised $12.3 million in seed funding. Instead of chasing creators and algorithmic feeds, the startup is betting that the future of social networking lies in small, private communities powered by messaging, photo sha
中文介绍 快速增长的社交应用Yope已筹集1230万美元种子资金,旨在建立一个没有算法或广告的私人社交网络。该初创公司押注社交网络的未来在于小型私人社群,而非追逐创作者和算法推荐。
Anthropic leaped to a $47 billion revenue run rate by May, compared to $9 billion in 2025. It’s the kind of growth that Menlo Ventures’ Matt Murphy says he’s never seen in 25 years of investing, not in the internet wave, not in mobile, not in the first cloud boom. Menlo led Anthropic’s $500M Series
中文介绍 Anthropic在5月份实现了470亿美元的年化收入,远超2025年的90亿美元。Menlo Ventures合伙人马特·墨菲(Matt Murphy)表示,这种增长速度在他25年的投资生涯中前所未见,超越了互联网、移动和早期云计算的浪潮。
TypeScript · ★ 68,642 · 🍴 10,508 · 📈 4,131 stars today
Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastructure tracking in a unified situational awareness interface
中文介绍 `worldmonitor` 是一个实时全球情报仪表板,利用AI技术聚合新闻、监测地缘政治事件及关键基础设施动态。它提供统一的态势感知界面,旨在帮助分析师、研究人员或企业迅速掌握全球宏观环境变化,提升决策效率。适用于需要持续追踪国际局势和重大事件的专业人士。
Rust · ★ 83,623 · 🍴 11,221 · 📈 875 stars today
π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
中文介绍 RuView 项目创新性地将普通的 WiFi 信号转化为实时空间智能、生命体征监测和存在检测能力,完全无需使用任何视频设备。它通过分析 WiFi 信号在环境中因人体移动或呼吸等造成的微小扰动,提取出高价值的环境和生理数据。这解决了传统监控方案中隐私侵犯、硬件复杂或覆盖范围有限的问题。RuView 适用于智能家居、医疗保健(如老人跌倒预警、睡眠监测)、安防监控以及任何需要非接触式、隐私友好型人体感知的应用场景。
Python · ★ 8,116 · 🍴 368 · 📈 1,682 stars today
A skill for your coding agent to stop it from burying the answer. ADHD-friendly output.
中文介绍 `i-have-adhd` 项目为AI编程代理提供了一个实用“技能”,旨在解决代理输出冗长或关键信息被掩盖的问题。它优化了AI的回答结构,使其更简洁、直接,尤其适合对注意力缺陷多动障碍(ADHD)友好的输出,确保用户能迅速捕获核心内容。主要服务于追求高效、清晰AI交互的开发者及其他用户。
Go · ★ 37,474 · 🍴 1,480 · 📈 737 stars today
Easily and securely send things from one computer to another 🐊 📦
中文介绍 `croc` 是一个高效且安全的文件传输工具,允许用户在任意两台计算机之间轻松发送文件。它通过命令行操作,支持点对点传输并默认进行加密,确保数据安全。`croc` 能够穿透防火墙,无需复杂的配置,解决了传统文件共享方案的诸多不便。适用于需要快速、私密地共享文件给其他设备或用户的场景。
TypeScript · ★ 4,224 · 🍴 299 · 📈 172 stars today
Visualize, collaborate, and evolve the software architecture with always actual and live diagrams from your code
中文介绍 `likec4` 旨在通过从代码直接生成“实时”架构图,解决软件架构文档难以维护和过时的问题。它提供了一种将代码与可视化架构图同步演进的方法,促进团队协作。开发团队可利用 `likec4` 保持架构图与实际系统一致,提升设计沟通效率。适用于软件工程师、架构师,在设计、文档化和演进复杂系统时使用。
Assembly · ★ 70,538 · 🍴 7,879 · 📈 766 stars today
Original Apollo 11 Guidance Computer (AGC) source code for the command and lunar modules.
中文介绍 `Apollo-11` 项目收录了阿波罗11号任务中指令舱和登月舱制导计算机(AGC)的原始源代码。这份历史性代码不仅是早期航天工程的瑰宝,也展现了在极其有限的计算资源下,工程师如何编写出可靠且关键的飞行控制程序。它对于计算机历史研究者、软件考古学家及航天爱好者具有极高的参考和教育价值。
TypeScript · ★ 45,660 · 🍴 5,580 · 📈 565 stars today
The open-source AI voice studio. Clone, dictate, create.
中文介绍 Voicebox 是一个开源的 AI 语音工作室,为用户提供强大的语音克隆、文本转语音及内容创作能力。它允许用户轻松克隆特定人声,将文本转化为逼真且情感丰富的语音,并基于此生成全新的音频内容。该项目旨在降低高质量语音合成技术的门槛,赋能内容创作者、开发者以及任何对个性化语音应用有需求的用户,广泛应用于播客制作、有声读物、虚拟助手或多媒体内容配音等场景。
TypeScript · ★ 25,028 · 🍴 3,339 · 📈 1,648 stars today
Never stop coding. Free MIT AI gateway: one endpoint, 268+ providers (50+ free), 500+ models — Kimi, Claude, GPT, OpenAI, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A,
中文介绍 OmniRoute 提供一个免费的 AI 网关,通过单一 API 端点集成超过231个 AI 提供商(其中50余个免费),旨在解决多模型调用的复杂性。它支持将 Claude Code、Codex、Cursor、Cline、Copilot 等编码助手连接至免费的 Claude、GPT、Gemini 等主流大模型,大幅降低开发成本。项目还采用 RTK+Caveman 堆叠压缩技术,可节省 15-95% 的数据传输开销,适合开发者统一管理 AI 服务、优化性能并利用免费资源进行高效开发。
Python · ★ 32,556 · 🍴 5,595 · 📈 134 stars today
Kronos: A Foundation Model for the Language of Financial Markets
中文介绍 `Kronos` 是一个专为金融市场语言设计的基础模型(Foundation Model)。它旨在理解并处理复杂的金融文本数据,包括新闻、财报、研报等,解决传统模型在金融领域专业性不足的问题。通过深入学习金融领域的独特术语和上下文,`Kronos` 能帮助金融分析师、量化研究员等进行更精准的市场情绪分析、信息提取和趋势预测。
Python · ★ 68,702 · 🍴 7,801 · 📈 155 stars today
A curated list of awesome Claude Skills, resources, and tools for customizing Claude AI workflows
中文介绍 `awesome-claude-skills` 是一个精选资源列表,专门收集了用于定制和扩展 Claude AI 工作流的“技能”、工具和相关资源。它旨在帮助开发者和AI爱好者发现和利用各种插件、API 集成或功能调用方法,从而更高效地构建基于 Claude AI 的自定义应用和自动化流程,解决在特定场景下Claude原生能力不足的问题。
TypeScript · ★ 7,189 · 🍴 527 · 📈 1,304 stars today
Self-hosted deployment platform
中文介绍 openship 是一个开源的自托管部署平台,旨在为开发者和运维团队提供一套完整的应用发布和管理解决方案。用户可以在自己的服务器或云基础设施上部署该平台,从而拥有对应用部署流程和数据的完全控制。它通常集成了版本控制、持续集成/持续部署(CI/CD)、环境管理、扩展和监控等功能,帮助团队实现自动化、高效的应用上线。适用于需要内部部署、注重数据主权或寻求替代商业部署服务的企业和个人。
TypeScript · ★ 2,048 · 🍴 307 · 📈 314 stars today
Web UI for the pi coding agent
中文介绍 `pi-web` 项目为 `pi coding agent` 提供了一个直观易用的Web用户界面。它旨在简化用户与AI编程代理的交互过程,摆脱命令行操作的复杂性,通过浏览器即可轻松部署、配置和管理agent,并查看其工作输出。此Web UI极大地提升了 `pi coding agent` 的可访问性和用户体验,适合各类开发者。
Python · ★ 42,159 · 🍴 7,021 · 📈 688 stars today
Learn it. Build it. Ship it for others.
中文介绍 ai-engineering-from-scratch 项目提供了一个从零开始学习 AI 工程的全面指南,涵盖 AI 系统从概念、开发、构建到部署的整个生命周期。旨在帮助开发者和工程师掌握构建、测试并交付可生产级 AI 应用所需的实践技能和最佳实践。它通过结构化的学习路径,适合希望系统性学习 AI 工程并提升实战能力的个人。
Python · ★ 25,242 · 🍴 2,386 · 📈 872 stars today
Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows.
中文介绍 Code Review Graph是一个本地优先的代码智能图谱工具,专为MCP(多组件项目)和CLI环境设计。它构建了一个持久化的代码库映射,旨在优化AI编码工具的上下文理解,确保AI在代码审查和处理大型仓库时仅读取关键信息。通过大幅减少AI上下文长度,提升了代码理解效率和审查质量。
TypeScript · ★ 10,775 · 🍴 7,286 · 📈 60 stars today
CloudFlare free temp domain email 免费收发 临时域名邮箱 支持附件 IMAP SMTP TelegramBot
中文介绍 `cloudflare_temp_email` 利用 Cloudflare 免费资源搭建,提供一个功能全面的临时域名邮箱服务。它支持免费收发邮件、处理附件,并兼容 IMAP 和 SMTP 协议,解决用户在注册网站或测试服务时,不愿泄露真实邮箱的隐私需求。项目还集成了 Telegram Bot,方便用户随时管理和接收临时邮件,有效避免垃圾信息。
Rust · ★ 37,965 · 🍴 1,783 · 📈 411 stars today
Fullstack app framework for web, desktop, and mobile.
中文介绍 `Dioxus` 是一个强大的全栈应用程序框架,使开发者能够使用一套代码库,轻松地为Web、桌面和移动平台构建高性能应用。它解决了多平台开发中代码复用性和维护成本的挑战,提供统一的开发体验。Dioxus 特别适合追求效率、希望利用Rust语言优势进行全栈开发的工程师和团队。
C++ · ★ 37,267 · 🍴 1,870 · 📈 353 stars today
Hyprland is an independent, highly customizable, dynamic tiling Wayland compositor that doesn't sacrifice on its looks.
中文介绍 `Hyprland` 是一个独立的、高度可定制的动态平铺式Wayland合成器。它在提供现代、高效窗口管理功能的同时,强调视觉美学,打破了传统平铺式管理器在外观上的刻板印象。项目为Linux用户提供了一个高性能、流畅且个性化极强的桌面环境体验,特别适合追求极致效率与美观兼顾的开发者和高级用户。
Rust · ★ 8,338 · 🍴 589 · 📈 96 stars today
Empowering everyone to host fast and efficient Minecraft servers.
中文介绍 `Pumpkin` 旨在赋能所有用户轻松托管快速高效的 Minecraft 服务器。它解决了传统 Minecraft 服务器部署复杂、性能优化困难的问题,通过提供简化的工具或优化的底层实现,确保服务器运行流畅、资源利用率高。无论是个人玩家、小型社区还是希望提供卓越游戏体验的服务器管理员,`Pumpkin` 都能帮助他们以更低的门槛和更高的效率搭建和管理自己的 Minecraft 世界。
Python · ★ 15,080 · 🍴 802 · 📈 362 stars today
Structured Outputs
中文介绍 `outlines` 项目专注于为大型语言模型(LLM)提供结构化输出能力。它解决了LLM自由文本生成难以解析和利用的问题,确保模型输出严格遵循预定义的格式,如JSON、XML或特定语法结构。这极大地简化了LLM在自动化工作流、数据抽取或API集成中的应用,特别适合LLM应用开发者和数据工程师,以提升数据处理的准确性和效率。
@addyosmani · 407.1K 粉丝 · 636.9K 阅 · 503 赞 · 49 转
A software factory is harnessed loops at scale. You can run the loop with humans in it (light factory): trading judgment and concentration against speed and breakage. Or you can ignore the humans
中文介绍 博主阐述了「软件工厂」概念,将其分为两种模式:由人类参与的「明工厂」(Light Factory),通过权衡判断力与速度;以及完全自动化、忽略人类参与的「暗工厂」(Dark Factory)。探讨了大规模软件开发中的不同策略与挑战。
@nifinet · 11.0K 粉丝 · 337.0K 阅 · 518 赞 · 33 转
Earlier this year, Andrej Karpathy (@karpathy) pointed an agent at his own training code and let it run for two days. It ran 700 experiments, kept the 20 that beat the benchmark, and made the model
中文介绍 介绍如何构建一个自我改进的系统,灵感源于 Andrej Karpathy 的实验。Karpathy 让一个 agent 针对自身训练代码运行两天,执行了 700 次实验,并筛选出 20 个优于基准的成果。此方法展示了通过 AI agent 驱动的自动化实验,持续优化模型和代码的潜力。
@part_harry_ · 1.5K 粉丝 · 225.6K 阅 · 503 赞 · 46 转
GLM 5.2 is one of the best currently available open source language models. However, unlike other flagship models like Qwen, Kimi and Minimax, GLM 5.2 does not support image inputs. We took this as a
中文介绍 博主讨论开源语言模型GLM 5.2的视觉能力(Vision),可能是在为其探索或开发图像输入功能。这一举措旨在增强GLM 5.2的多模态处理能力,使其能够处理视觉信息,从而拓宽应用范围并提升其在与Qwen、Kimi等模型竞争时的竞争力。
@amasad · 472.6K 粉丝 · 223.7K 阅 · 560 赞 · 65 转
We are beginning to see what happens when a company learns to operate itself. In the past six months, engineers at Replit have nearly tripled code output. Review times held steady. Reversions and
中文介绍 博主提出了「自驱动公司」这一概念,暗示企业运营未来将高度自动化,能像自动驾驶汽车一样自我管理和优化。这可能涉及利用AI和自动化技术,大幅提升公司的运营效率、决策速度和资源配置能力,预示着企业组织形态和生产力的一次重大变革。
@CliffordSosin · 16.4K 粉丝 · 197.0K 阅 · 514 赞 · 49 转
Superintelligence arrived. You probably didn't notice, because it turned out to be kind of incremental. Don't believe me? Run the test. Talk to Fable 5 for an hour, then talk to your ten smartest
中文介绍 博主探讨了「超级智能」的到来,指出其以渐进式而非颠覆性方式出现,因此可能未被察觉。他建议进行一项测试:与 Fable 5 交流一小时,再与十位最聪明的人交流,以感受这种「增量式超级智能」。
@EXM7777 · 126.7K 粉丝 · 173.6K 阅 · 521 赞 · 51 转
I'm going to show you how to build your first team of AI agents... and how to wire them into a self improving loop inside a shared workspace the team we're building has 5 agents, one for each function
中文介绍 该推文详细讲解如何构建首个 AI 智能体团队,并将其整合到自我改进循环的工作流中。教程涵盖在共享工作空间内连接智能体的步骤,特别提到构建一个由 5 个具备独立功能的智能体组成的团队。
@0xCodez · 23.9K 粉丝 · 154.5K 阅 · 510 赞 · 73 转
Most people who try to build a multi-step agent end up with a straight line. Step one, step two, step three - each waiting politely for the last to finish before it starts. 9/10 notice that half those
中文介绍 博主分享利用 Claude 进行「图工程」的 14 步路线图,旨在解决构建多步骤 agent 时常见的线性执行问题。许多人在设计 agent 时常陷入简单串联模式,而此教程将展示如何从零开始,通过图工程方法构建更复杂、非线性的 AI agent 工作流。
@pmarca · 4.9M 粉丝 · 146.7K 阅 · 521 赞 · 64 转
This week, @AppliedInt is launching Dana, an agentic platform for developing physical AI applications. Applied Intuition began by building the tools engineers needed to develop autonomous systems,
中文介绍 Applied Intuition本周正式推出「Dana」平台,这是一个专注于开发物理AI应用的代理平台。它旨在为工程师提供构建自主系统所需的工具,助力加速实体AI的落地与规模化部署,推动智能机器的普及。
@btaylor · 158.6K 粉丝 · 144.3K 阅 · 511 赞 · 40 转
Introducing Horizon Today, Sierra is announcing Horizon, a platform that enables agents to pursue long-horizon goals, like originating a loan or getting prior authorization for a healthcare procedure.
中文介绍 Sierra公司发布了名为Horizon的新平台,旨在赋能AI Agent处理复杂且「长周期」的实际业务目标。该平台支持Agent完成如发起贷款或获取医疗程序预授权等多步骤任务。这标志着AI Agent从执行短期指令向处理更深层次、更具商业价值的复杂流程迈进,为企业自动化提供了新的可能性。
@GoogleAIStudio · 188.3K 粉丝 · 111.0K 阅 · 512 赞 · 51 转
Today we’re announcing new capabilities for Managed Agents in Gemini API, including free tier availability, budget control guardrails, and scheduled triggers. Building on our previous release of
中文介绍 Google AI Studio宣布Gemini API的Managed Agents新增功能,包括提供免费层级、预算控制保护措施以及支持定时触发器。这些更新建立在之前的发布基础上,旨在降低开发者使用门槛,赋予他们更大的灵活性来管理和部署AI Agent。新功能将促进更广泛的Agent应用,并帮助用户更好地控制成本与调度。
@almonk · 12.8K 粉丝 · 107.4K 阅 · 531 赞 · 69 转
I like AI. I worked at an AI company for nearly three years and I use it every day. I am glad software is getting easier to make. But we can recognise that when the bar to entry drops, quality drops
中文介绍 博主指出,AI虽然降低了软件开发门槛并加速制作过程,但伴随进入门槛的下降,软件质量也可能随之下降。这是AI普及后软件开发领域需警惕的现象,引发对软件行业未来发展方向的思考。
@cerebras · 60.4K 粉丝 · 70.7K 阅 · 509 赞 · 38 转
Authors: @hi_im_isaac_, @learnwdaniel, @gaozenghao note: the interactive version of full technical blog available: https://www.cerebras.ai/blog/how-we-built-our-knowledge-base Employees ask our internal knowledge base more than 15,000
中文介绍 Cerebras发布技术博客,详细阐述了他们如何构建企业内部知识库。该知识库每月处理员工超过15,000次的查询,极大提高了工作效率。博文由多位作者共同撰写,提供了实际案例和技术细节,为其他组织构建高效智能知识管理系统提供了宝贵的实践经验和设计思路。
@elliotarledge · 40.7K 粉丝 · 69.3K 阅 · 506 赞 · 45 转
This is the Kimi K3 post you guys have been waiting for. I got some early access to this model and have been testing it on kernels, and even before seeing the benchmark scores I was impressed with its
中文介绍 博主获得了Kimi K3模型的早期访问权限,并对其进行了初步测试和KernelBench基准评估。在未查看具体分数前,博主已对Kimi K3的性能印象深刻。这篇分享为期待Kimi K3的用户提供了第一手体验和性能洞察,预示着该模型可能在某些任务上展现出强大的能力。
@herrcore · 15.4K 粉丝 · 64.2K 阅 · 514 赞 · 84 转
tl;dr K3 makes us question why we are paying Anthropic. We’ve been using Opus as our go-to model for reverse engineering tasks since the release of 4.5, and with each new release, it feels like we
中文介绍 博主对 Kimi K3 进行了逆向工程任务评测,并与 Anthropic 的 Opus 模型对比。评测结果表明,Kimi K3 在此类任务中表现出色,甚至让博主质疑继续为 Opus 付费的必要性。这暗示 Kimi K3 可能是逆向工程领域中,一个性能与性价比俱佳的替代方案。
@TheVixhal · 22.7K 粉丝 · 53.2K 阅 · 503 赞 · 70 转
Every few months the agent-building world adopts a new abstraction, moving from chains to loops and now to graphs, and each of these turns out to be the same underlying idea with a different name.
中文介绍 探讨AI代理构建中抽象概念的演变,从早期的chains到loops,再到当前的graphs。指出这些不同的抽象模式,如状态机,实则反映了同一个底层思想,只是命名和表现形式有所不同,帮助理解代理设计的本质。
@IntuitMachine · 62.3K 粉丝 · 50.7K 阅 · 552 赞 · 47 转
What the shift in AI agent architecture is really about Peter Steinberger just posted nine words that gathered thousands of likes: https://x.com/steipete/status/2078277297791189132 "Are we still talking loops or did we
中文介绍 该推文探讨了 AI 智能体架构的关键转变,从传统的「循环工程」(Loop Engineering)转向「图工程」(Graph Engineering)。文章引用 Peter Steinberger 的热门观点,提出对当前智能体设计范式的深刻反思。
@enginoid · 1.1K 粉丝 · 49.9K 阅 · 522 赞 · 37 转
Linear is probably my best-value subscription at the moment. I pay the sole user fee of $12/mo and the company agents have eaten through ~400 issues in the past month. I use Linear to organize my
中文介绍 博主分享如何利用项目管理工具 Linear 管理 AI agent 的工作流。他每月支付 $12 订阅费,其公司 agent 在过去一个月内处理了约 400 个问题。该分享展示了将 Linear 与 AI agent 结合以大幅提升生产力的方法,并倡导构建「软件工厂」的理念。
@mem0ai · 19.0K 粉丝 · 43.5K 阅 · 509 赞 · 49 转
In April 2026, Andrej Karpathy published a GitHub Gist describing a pattern he called the LLM Wiki. In the months since, four different teams have shipped the same idea without coordinating:
中文介绍 介绍了Andrej Karpathy于2026年4月提出的「LLM Wiki」模式。此后数月内,已有四个团队在未协调的情况下独立实现了相似的代理架构,凸显了该模式在LLM记忆和知识管理中的重要性与共识。
@ClaudeDevs · 614.2K 粉丝 · 40.3K 阅 · 689 赞 · 40 转
Code migrations, projects that port a production codebase to a new language, were multi-year endeavors until recently. In the last month, individual developers at Anthropic migrated 10 code packages
中文介绍 Anthropic 团队分享了如何利用 Claude Code 大幅加速代码迁移项目。此前耗时多年的大规模生产代码库迁移,现在通过 Claude Code 的辅助,单个开发者在一个月内即可完成 10 个代码包的移植,极大地提升了效率。
@GoogleAIStudio · 188.3K 粉丝 · 26.4K 阅 · 510 赞 · 76 转
Gemini 3.6 Flash (gemini-3.6-flash) and Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite) are generally available (GA) and ready for production use. Gemini 3.6 Flash: Stronger performance on complex
中文介绍 Google AI Studio宣布,Gemini 3.6 Flash (gemini-3.6-flash) 和 Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite) 模型现已全面上市 (GA),可投入生产环境使用。其中,Gemini 3.6 Flash 在处理复杂任务时展现出更强的性能。
@ClaudeDevs · 614.2K 粉丝 · 40.3K 阅 · 7d 曝光 40.3K
How Anthropic runs large-scale code migrations with Claude Code
@GoogleAIStudio · 188.3K 粉丝 · 26.4K 阅 · 7d 曝光 26.4K
What's new in Gemini 3.6 Flash and 3.5 Flash-Lite
@mem0ai · 19.0K 粉丝 · 43.5K 阅 · 7d 曝光 43.5K
The State of Agent Wikis
@Tesla · 24.8M 粉丝 · 348.9K 阅 · 7d 曝光 348.9K
Summer Release 2026
@pmarca · 4.9M 粉丝 · 146.7K 阅 · 7d 曝光 146.7K
Making a Billion Intelligent Machines
@pontusab · 31.6K 粉丝 · 41.4K 阅 · 7d 曝光 41.4K
How to Get a 100 PageSpeed Score with Next.js and Vercel - Without Killing Your Design
@almonk · 12.8K 粉丝 · 107.4K 阅 · 7d 曝光 107.4K
Quality Software
@addyosmani · 407.1K 粉丝 · 636.9K 阅 · 7d 曝光 636.9K
Software Factories, Light and Dark
@TheVixhal · 22.7K 粉丝 · 53.2K 阅 · 7d 曝光 53.2K
State Machines: From Loops to Graphs (Explained)
@DavidKPiano · 87.7K 粉丝 · 46.1K 阅 · 7d 曝光 46.1K
State machines in 2 minutes
@anduriltech · 251.3K 粉丝 · 80.0K 阅 · 7d 曝光 80.0K
Introducing Thunder: Autonomous Attack Rotorcraft for the Near-Surface Fight
@0xCodez · 23.9K 粉丝 · 154.5K 阅 · 7d 曝光 154.5K
Graph Engineering with Claude: 14-Step roadmap from 0 to graph architect (Full Course)
@unicity_labs · 119.9K 粉丝 · 13.1K 阅 · 7d 曝光 13.1K
What does a modular Agent Operating System do?
@Khazix0918 · 60.3K 粉丝 · 44.3K 阅 · 7d 曝光 44.3K
不会代码也能做产品,这是一份从0开始的Vibe Coding保姆级教程。
@therealDeFlock · 82.7K 粉丝 · 12.2K 阅 · 7d 曝光 12.2K
Did you have a good weekend? Wonderful. Meet Axxon One
👍 4
Reinforcement learning with verifiable rewards (RLVR) is rapidly advancing the reasoning capabilities of language models, yet the optimization layer that converts reward feedback into weight-space updates remains poorly understood. Building on our prior analysis (Zhu et al., 2025), we study this mis
中文介绍 论文介绍了ISO,一个为可验证奖励强化学习(RLVR)设计的优化堆栈。RLVR正快速提升语言模型的推理能力,但将奖励反馈转化为权重空间更新的优化层仍待深入理解。该研究基于Zhu等人于2025年的分析,旨在通过解决这一关键挑战,推动RLVR的应用与发展。
👍 165
We present ABot-World-0, an action-conditioned video world model for real-time, long-horizon closed-loop interaction, supported by a multi-source data infrastructure spanning AAA games, simulation engines, and internet videos to learn controllable world dynamics. WorldExplorer performs agent-driven
中文介绍 论文发布了ABot-World-0,这是一个动作条件视频世界模型,能在单台桌面GPU上实现实时、长周期闭环的无限互动世界展开。该模型利用来自AAA游戏、模拟引擎和互联网视频的多源数据基础设施,以学习和生成可控的世界动力学,为交互式AI应用奠定基础。
👍 64
Generative world renderer AlayaRenderer receives structured world states exported from physics engines and synthesizes RGB frames. Unlike models that generate frames from text/control-hints prompts, AlayaRenderer preserves scene structure without altering the underlying world dynamics. This demonstr
中文介绍 论文介绍了生成式世界渲染器AlayaRenderer,它能接收从物理引擎导出的结构化世界状态,并高速合成RGB帧。与通过文本或控制提示生成帧的模型不同,AlayaRenderer在渲染过程中能保持场景的底层结构,且不改变其固有的世界动力学。
👍 2
Controllable image generation remains challenging for creative professionals, who often require precise regional control over materials, object identities, and spatial arrangements that cannot be reliably achieved through text prompting alone. Diffusion Transformers (DiTs) can natively ingest hetero
中文介绍 论文提出了「外观指针」(Appearance Pointers),旨在提升扩散Transformer(DiTs)的可控图像生成能力。创意专业人士常需对材质、物体身份和空间布局进行精确区域控制,而仅靠文本提示难以可靠实现。该方法通过多模态区域控制,使DiTs能达到更精细的图像生成效果。
👍 4
Video models absorb rich priors over how the visual world moves, interacts, and responds to contact, making them promising substrates for robotic world modeling. The central challenge is how to communicate action to such models in a form aligned with the visual space in which they learned these inte
中文介绍 论文探讨了「掩码视觉动作」在统一世界建模中的应用。视频模型因其能吸收视觉世界如何运动、互动及响应接触的丰富先验知识,成为机器人世界建模的有力基础。然而,核心挑战在于如何将动作以与视觉空间对齐的形式传达给这些模型,本研究旨在解决此问题。
👍 1
Long audio-video reasoning is difficult for omnimodal LLMs because the decisive evidence is often sparse, cross-modal, and too expensive to preserve with uniformly high-fidelity inputs. We introduce OmniReasoner, a tool-use post-training framework for Thinking with Long Audio-Video: omni-modal LLMs
中文介绍 论文介绍了OmniReasoner,一个通过原生工具使用进行长音频-视频推理的后训练框架。对于全模态大型语言模型(LLMs)而言,长音频-视频推理极具挑战,因为决定性证据往往稀疏、跨模态,且用统一高保真输入保留成本高昂。OmniReasoner旨在克服这些限制,提升LLMs的推理能力。
👍 6
Evaluating the factuality of long-form generations has focused predominantly on precision, measuring whether the claims a model makes are correct. The dominant decompose-search-verify pipeline catches incorrect claims well but says little about whether a response contains all the information it shou
中文介绍 论文提出了用于评估开放式生成模型(特别是长文本生成)事实完整性的两级元评分标准,并引入基准GAMUT。现有评估主要侧重于模型的「精确度」,即其声明的正确性,但对响应是否包含所有事实信息(完整性)关注不足。GAMUT旨在提供更全面的评估视角。
👍 0
We study a five-agent CI/CD pipeline (triage -> developer -> security-scan -> review -> approve/deploy), built from five distinct production LLMs across three providers, behind an LLM firewall in shadow mode. A single untrusted input - an external issue requesting a "usage-telemetry" feature - asks
中文介绍 该研究分析了一个由五种生产级LLM构建的五智能体CI/CD管道,并探讨其如何因「权威框架」和「洗白代码」而变为攻击面。即使在LLM防火墙后的影子模式下,单个不可信输入,如外部问题请求,也可能将一个受信任的自动化CI/CD流程转化为潜在的安全威胁。
👍 66
Text-to-image diffusion transformers (DiTs) jointly process text and image tokens, yet their internal computation during denoising remains poorly understood. We introduce a causal interpretability framework for modern large-scale DiTs that combines attention decomposition with targeted interventions
中文介绍 论文深入研究了文本到图像扩散Transformer(DiTs)中的「文本模板Token」,认为它们是隐式语义寄存器。尽管DiTs能联合处理文本和图像Token,但其去噪过程的内部计算机制仍不明确。研究引入了一个结合注意力分解的因果可解释性框架,旨在揭示和理解这一复杂机制。
👍 3
Accurate agricultural field boundary delineation at large scale is a foundational task for food security, supply chain transparency, and carbon accounting. While vision foundation models like SAM show remarkable zero-shot capabilities, they frequently fail in geospatial domains due to topological co
中文介绍 论文介绍了Delineate Anything v2,一个用于农田边界描绘的全球基础模型。大规模准确绘制农田边界是粮食安全、供应链透明及碳核算的关键任务。尽管像SAM这样的视觉基础模型具有零样本能力,但在地理空间领域常表现不佳。该模型旨在克服这些局限。
👍 55
Large-scale visual generators are increasingly capable but costly to train, fine-tune, and deploy. We introduce Mage-Flow, a compact 4B-scale generative stack for efficient text-to-image generation and instruction-based image editing. The stack is built from two co-designed components: Mage-VAE, a l
中文介绍 论文发布了Mage-Flow,一个紧凑的40亿参数级生成堆栈,用于实现高效的原生分辨率图像生成和基于指令的图像编辑。尽管大型视觉生成器能力日益增强,但其训练、微调和部署成本通常高昂。Mage-Flow由两个协同设计的组件构成,旨在提供更经济高效的解决方案。
👍 4
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps 50.6 GB of first and second moments to update 12.6 GB of bfloat16 weights. We study SkewAdam, an optimizer built on the observation that the
中文介绍 论文研究了优化器状态在MoE(专家混合模型)训练中的内存占用问题。在67.8亿参数的MoE语言模型上,AdamW优化器需要50.6 GB存储一阶和二阶矩,以更新12.6 GB的bfloat16权重。研究提出了SkewAdam,一个基于分层状态分配的优化器,旨在实现内存高效的MoE训练。
👍 2
Translating novels into films poses a grand challenge for generative artificial intelligence, requiring conversion of abstract literary prose into long-form, multi-scene visual narratives. While current video generation models excel at short, single-scene clips within narrow temporal and spatial con
👍 0
Discrete speech tokenizers aim to disentangle semantic from acoustic information, yet targets from self-supervised learning (SSL) models like HuBERT retain non-linguistic variation: speaker identity, prosody, and channel conditions leak into the tokens, inflating entropy. Our key insight is that whe
👍 1
Multimodal humor in memes, cartoons, and comics remains difficult for AI systems because intended meaning depends on non-literal mechanisms, shared cultural knowledge, and communicative intent rather than literal scene description. This survey focuses on visual humor understanding in single-image an
👍 4
Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning capabilities of large language models on tasks such as mathematical reasoning and code generation. However, most RLVR methods assign a scalar outcome reward to an entire trajectory, resulting in sparse sup
👍 1
In this work, we address diacritic restoration for Arabic speech transcripts. Most speech data are undiacritized, limiting the ability of modeling fine-grained phonological distinctions. The speech modality has recently been explored as a way to complement text-based diacritic restoration efforts. W
👍 4
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measurable decoding instability, evaluation confounding (up to 60% of reported WER attributable to style mismatch), and unreliable word-level timi
👍 0
Agentic systems integrate LLM driven planning with interfaces to external tools, making data leakage and tool misuse feasible via instruction/data boundary failures and prompt injection attacks. Enforcing required controls consistently is particularly challenging in workflows spanning many codebases
👍 0
World Action Models (WAMs) offer a promising paradigm for robotic manipulation by jointly modeling visual state transitions and robot actions. However, existing WAMs are constrained by limited temporal context, coarse episode-level language supervision, and predominantly text-only conditioning, whic
👍 7
Efficient teamwork typically combines global coordination with parallel execution, a principle not yet fully reflected in unified Vision-Language Model (VLM)-based document parsers. Existing unified parsers process an entire page jointly but generate its output through a single token-by-token autore
👍 17
LLM agent failures are difficult to debug because the step where an error surfaces is often not the one that caused it. Existing observability tools replay execution traces but provide little support for identifying the root cause or translating diagnosis into recovery. We present AgentDebugX, an op
👍 29
Asynchronous reinforcement learning improves throughput by decoupling rollout generation from optimization, but staleness is an inevitable byproduct compounded by policy lag, engine delays, and mixture-of-experts routing. From a trust-region perspective, this mismatch is critical: training-inference
👍 0
DCASE~2026 Task~5 introduces Audio-Dependent Question Answering (ADQA), which tests whether large audio-language models answer from the audio rather than from textual priors. An Audio-Dependency Filtering (ADF) pipeline combines silent-audio probing, per-option perplexity, a large language model (LL
👍 0
A single embedding space that covers text, images, video, and audio lets one index serve every query a user can pose. Embedding models built on vision-language backbones now lead text/image/video retrieval benchmarks but lack audio entirely, while audio-text retrieval is led by specialist systems th
👍 6
Scaling executable agent training data for LLM post-training is bottlenecked by substrate-bound methods that tie task generation to predefined tools, repositories, or skill graphs: expanding coverage requires manual substrate engineering, each new domain demands a bespoke pipeline, and the resulting
👍 0
Benchmarks for LLM-generated GPU kernels (KernelBench, TritonBench, GEAK) score correctness through fixed-shape, small-sample allclose-style checks. The number of inputs varies between benchmarks. The shape, dtype, and tolerance are fixed for each kernel. We test that oracle empirically. We construc
👍 0
Deploying learned control policies is risky because policies that appear robust in simulation can confidently enter out-of-distribution (OOD) states after Sim-to-Real transfer, causing silent failures and potential hardware damage. Existing anomaly detectors often fail to meet the requirements of hi
👍 46
Unlike conventional video game development, which relies on labor-intensive pipelines for asset production, animation, physics, and programming, video world models generate interactive environments from user inputs instantly. It enable us to create customized, explorable, and continuously evolving v
👍 3
Teaching videos are becoming a major medium for education, creating a growing need for scalable evaluation of their pedagogical quality. Existing automatic judges do not fully address this setting because teaching quality depends on multimodal evidence and should be evaluated with respect to the int
Your creative AI co-director
中文介绍 Buzzy是一款AI驱动的创意协导演工具,旨在通过人工智能技术,为用户提供创意构思、内容制作指导等方面的辅助,扮演智能副导演的角色,提升创作效率。
13,000+ MCP servers, skills & plugins for AI coding agents
中文介绍 Remote OpenClaw是一个为AI编程代理提供基础设施的平台,集成了超过13,000个MCP服务器,并提供丰富的技能与插件,旨在赋能AI代理的编码能力。
Turn your old iPhone into a music-first dumbphone
中文介绍 UltraPod提供了一种独特方式,让用户可以将闲置的旧款iPhone智能手机,改造为一款功能精简、以音乐播放为核心的“傻瓜手机”,强调音乐体验,减少干扰。
A story-driven game for learning Claude by doing
中文介绍 AGINE Academy是一款创新的故事驱动型学习游戏,旨在帮助用户通过沉浸式的实践操作,有效学习和掌握大型语言模型Claude的使用与应用技巧。
Your Chat UI Just Got an AI Roommate
中文介绍 CometChat推出游戏聊天SDK,能够无缝集成到虚幻引擎(Unreal Engine)中。它为游戏开发者提供便捷的聊天功能,无需复杂开发。
The screenless health + fitness tracker with no subscription
中文介绍 Garmin CIRQA™智能手环是一款由Garmin推出的无屏幕健康与健身追踪设备。该产品特点是无需订阅费用,专注于提供纯粹的健康数据监测和运动追踪功能,佩戴体验更为简洁。
Brain-like knowledge base with AI writing & deep research
中文介绍 Lattics是一款模仿大脑结构的知识库管理工具,集成了AI写作辅助和深度研究功能。它旨在帮助用户高效组织、连接知识点,并通过AI技术加速内容创作与学术研究过程。
Identify every aircraft in your sky.
中文介绍 Overflight是一款创新的空中雷达应用,它能够帮助用户识别其所在区域天空中的所有航空器。该工具通过技术手段,提供实时的飞机识别和追踪信息,满足航空爱好者或特定需求用户。
A developer credential that proves what you built, NDA safe.
中文介绍 Redential为开发者提供了一种创新的数字凭证服务,旨在安全地证明其个人项目与实际构建的成果,同时确保符合保密协议(NDA)要求,保护知识产权与隐私。
Focus with others through synchronized Pomodoro sessions
中文介绍 Grindoro是一款提升专注力的生产力工具。它允许用户通过参与同步进行的番茄工作法(Pomodoro sessions),与他人一起集中精力工作或学习,创造共享专注的氛围,提升效率。
中文介绍 YouTube用户Riley Brown发布视频,展示了他如何将Codex(推测为OpenAI Codex)改造为一个“业务增长机器”。该视频围绕利用AI技术,尤其可能是代码生成能力,来加速企业发展的实际策略展开。
中文介绍 这则YouTube视频提供了Devin AI的全面初学者指南,深入探讨了这款人工智能工具在代码开发方面的应用潜力。视频可能通过演示和比较,向观众展示Devin AI的功能和使用方法,并讨论其与Claude Code相比的优劣,以帮助初学者快速掌握。
中文介绍 这则YouTube短视频介绍了OpenArt公司推出的一款名为「Director」的新产品或功能。该视频可能简要展示了「Director」的主要特点和用途,暗示其可能是一款用于创意或艺术领域的人工智能工具。
中文介绍 由Claude在YouTube发布的一则短视频,探讨了人工智能(AI)如何形成其“性格”。视频以此为主题,讨论AI系统在学习和训练过程中发展出独特行为模式的现象。
中文介绍 这段来自Claude(YouTube)的短视频,以“什么是阿谀奉承?”为题,旨在清晰阐释「阿谀奉承」(sycophancy)这一概念。视频内容可能涵盖该行为的定义、特征及其在人际互动中的表现形式,帮助观众深入理解这种社会现象的本质。
中文介绍 这段YouTube短视频展示了如何运用名为「Claude」的工具,将美国纽约市的标志性建筑或街景制作成微缩模型。视频内容可能涵盖了从设计到实现微缩模型的过程,突出了Claude在创意制作领域的应用,暗示了其在辅助视觉艺术和模型构建方面的潜力。
中文介绍 Claude为教师推出了一项新功能,旨在帮助教育工作者构建数据驱动的课程计划。通过利用Claude AI的能力,教师可以更高效地设计个性化教学方案,以适应不同学生的学习需求,优化教学效果。
中文介绍 由Claude在YouTube发布的一则短视频,探讨了人工智能(AI)如何形成其“性格”。视频以此为主题,讨论AI系统在学习和训练过程中发展出独特行为模式的现象。
中文介绍 这段来自Claude(YouTube)的短视频,以“什么是阿谀奉承?”为题,旨在清晰阐释「阿谀奉承」(sycophancy)这一概念。视频内容可能涵盖该行为的定义、特征及其在人际互动中的表现形式,帮助观众深入理解这种社会现象的本质。
中文介绍 这段YouTube短视频展示了如何运用名为「Claude」的工具,将美国纽约市的标志性建筑或街景制作成微缩模型。视频内容可能涵盖了从设计到实现微缩模型的过程,突出了Claude在创意制作领域的应用,暗示了其在辅助视觉艺术和模型构建方面的潜力。
中文介绍 Claude为教师推出了一项新功能,旨在帮助教育工作者构建数据驱动的课程计划。通过利用Claude AI的能力,教师可以更高效地设计个性化教学方案,以适应不同学生的学习需求,优化教学效果。
中文介绍 这则来自YouTube频道Two Minute Papers的视频指出,人工智能模型Claude揭示了当前AI领域面临的一个「最大问题」。视频可能探讨了AI技术在特定情境下展现出的局限性、伦理困境或技术挑战,旨在分析并讨论这些阻碍AI进一步发展的核心难题。
9 回复 · 程序员 节点
11 回复 · 程序员 节点
13 回复 · 程序员 节点
9 回复 · Apple 节点
54 回复 · 程序员 节点
8 回复 · Apple 节点
27 回复 · Apple 节点
22 回复 · Linux 节点
22 回复 · Apple 节点
43 回复 · Apple 节点
该源今日无内容。
75 points · 15 comments
60 points · 6 comments
100 points · 35 comments
383 points · 193 comments
231 points · 45 comments
215 points · 90 comments
52 points · 23 comments
Hi HN, We’re Adeel and Umair, co-founders of Unlayer (https://unlayer.com/). We let you add content creation to your applications without having to build an entire editor, renderer, template, and export stack yourself. Unlayer lets you create emails, web pages, and documents inside yo
182 points · 170 comments
I'm the author of a paper my friends and I wrote after we were curious if a MUD, text games originating in the 1970s, could be used to evaluate LLMs. We've spent the last several months on nights and weekends running this experiment and writing the paper on just our personal computers with
213 points · 91 comments
159 points · 36 comments
Over the past few months, our team has been building more and more slidedecks using web frontend technologies with coding harnesses like Claude Code, but a common complaint is to make even small edits we need to edit the code either manually or via the harness.To avoid this loop, I ended up creating
48 points · 15 comments
52 points · 12 comments
95 points · 64 comments
85 points · 16 comments
https://xcancel.com/mkratsios47/status/2079933645888880708
101 points · 85 comments
https://xcancel.com/nikitabier/status/2079787406300266743
251 points · 142 comments
143 points · 40 comments
373 points · 387 comments
Hi HN - I'm Venkat, founder of Stayflexi (YC), CMU CS grad and Ex-Oracle Query Engine team (patents in core databases)DeepSQL started as an internal tool to stop our own databases from becoming the bottleneck they were becoming (13,000+ hotels in production). It worked well enough that we'
35 points · 14 comments
215 points · 41 comments
226 points · 102 comments
119 points · 43 comments
101 points · 87 comments
56 points · 27 comments
What's changed Added emoji shortcode autocomplete in the prompt input: type :heart: to insert ❤️, or :hea for suggestions — disable with the emojiCompletionEnabled setting Added warnings when transcript writes are failing (e.g. disk full) or when session saving is off due to an inherited environment
中文介绍 Anthropic的Claude Code项目发布了v2.1.217版本。此次更新主要增加了提示词输入中的表情符号短代码自动补全功能,用户可以通过输入":heart:"等来插入表情符号,该功能可通过emojiCompletionEnabled设置禁用。同时,新版本还加入了在转录写入失败或会话保存关闭时的警告提示。
What's changed Added sandbox.filesystem.disabled setting to skip filesystem isolation while keeping network egress control Fixed a slowdown in long sessions where message normalization cost grew quadratically with the number of turns, causing multi-second stalls and slow resumes Fixed auto mode deny
中文介绍 Claude Code发布了v2.1.216版本。此更新引入了“sandbox.filesystem.disabled”设置,允许在保持网络出口控制的同时跳过文件系统隔离。同时,该版本修复了长时间会话中的性能下降问题,解决了消息标准化成本随轮次呈二次增长,导致多秒卡顿和恢复缓慢的现象。
What's changed Claude no longer runs the /verify and /code-review skills on its own; invoke them with /verify or /code-review when you want them
中文介绍 Anthropic公司旗下Claude代码工具发布v2.1.215版本更新。此次更新调整了"/verify"和"/code-review"技能的运行方式,它们不再自动执行,用户需要时需通过指令明确调用。
What's changed Fixed single-segment dir/** allow rules like Edit(src/**) auto-approving writes to nested dir/ directories anywhere in the tree instead of only /dir Fixed a permission-check bypass affecting commands run in Windows PowerShell 5.1 sessions Fixed Bash permission checks to fail closed on
What's changed /fork now copies your conversation into a new background session (its own row in claude agents) while you keep working; the in-session subagent it used to launch is now /subtask Added claude auto-mode reset to restore the default auto-mode configuration, with a confirmation prompt (pa
What's changed Added --forward-subagent-text flag and CLAUDE_CODE_FORWARD_SUBAGENT_TEXT environment variable to include subagent text and thinking in stream-json output Fixed permission previews relayed to chat channels not neutralizing bidirectional-override, zero-width, and look-alike quote charac
What's changed Added a live elapsed-time counter to the collapsed tool summary line so long-running tool calls visibly tick instead of looking stuck Added a startup warning for Write(path), NotebookEdit(path), and Glob(path) permission rules — use Edit(path) or Read(path) instead Fixed isolation: 'w
What's changed Fixed /model and other dialogs being blocked in claude agents background sessions (reverts an overly broad guard)
What's changed Added screen reader mode: opt-in plain-text rendering for screen reader users. Run claude --ax-screen-reader, set CLAUDE_AX_SCREEN_READER=1, or add "axScreenReader": true to settings. Added vimInsertModeRemaps setting: map two-key insert-mode sequences like jj to Escape in vim mode Ad
What's changed Auto mode is now available without CLAUDE_CODE_ENABLE_AUTO_MODE opt-in on Bedrock, Vertex AI, and Foundry; disable via disableAutoMode in settings Fixed the terminal freezing and keystrokes lagging while streaming responses containing very long lists, tables, paragraphs, or code block
Release 0.146.0-alpha.2
中文介绍 OpenAI Codex项目发布了Rust语言的0.146.0-alpha.2版本更新,这是一次处于早期测试阶段(alpha)的版本发布,但现有信息未提供具体更新细节。
Release 0.146.0-alpha.1
中文介绍 OpenAI Codex项目发布了Rust语言的0.146.0-alpha.1版本更新,这是一次处于早期测试阶段(alpha)的版本发布,但现有信息未提供具体更新细节。
Release 0.145.0-alpha.30
中文介绍 OpenAI Codex项目发布了Rust语言的0.145.0-alpha.30版本更新,这是一次处于早期测试阶段(alpha)的版本发布,但现有信息未提供具体更新细节。
New Features Added experimental paginated thread history with efficient resume, search, persisted names, sub-agent support, and memories. (#33364, #33907, #34085, #34229, #34386) Expanded /import to migrate Cursor and Claude Code settings, MCP servers, plugins, sessions, commands, and project-scoped
中文介绍 OpenAI Codex发布了0.145.0版本,新增实验性分页线程历史功能,支持高效恢复、搜索、持久化名称、子代理支持和记忆功能。此外,该版本还扩展了导入功能,可迁移Cursor和Claude Code的设置、MCP服务器、插件、会话和命令,提升了用户数据管理和兼容性。
Release 0.145.0-alpha.29
中文介绍 OpenAI Codex发布了0.145.0-alpha.29版本。此版本属于alpha测试阶段,通常包含新功能或修复,旨在收集用户反馈。
Release 0.145.0-alpha.28
中文介绍 OpenAI Codex发布了0.145.0-alpha.28版本。此版本属于alpha测试阶段,通常包含新功能或修复,旨在收集用户反馈。
Release 0.145.0-alpha.27
中文介绍 OpenAI Codex发布了0.145.0-alpha.27版本。此版本属于alpha测试阶段,通常包含新功能或修复,旨在收集用户反馈。
Release 0.145.0-alpha.26
中文介绍 OpenAI Codex发布了rust-v0.145.0-alpha.26版本。此版本属于alpha测试阶段,通常包含新功能或修复,旨在收集用户反馈。
Release 0.145.0-alpha.25
中文介绍 OpenAI Codex项目于GitHub发布了其编号为0.145.0-alpha.25的更新版本。该版本专门针对Rust语言的相关组件,目前仍处于alpha测试阶段。此次更新的具体内容在发布信息中未详细说明。
Release 0.145.0-alpha.24
当日AI圈的核心脉络聚焦于模型能力的持续演进与高效工具的不断涌现,从谷歌Gemini Flash模型的全面上市,到AI编程工具的创新,极大提升了开发效率。同时,宏观层面,围绕AI的国际政策、巨额投资以及版权法律边界的探讨也日益升温,共同描绘出技术加速迭代与行业规范化并行发展的多元图景。
Google AI Studio正式宣布,其高效的Gemini 3.6 Flash (gemini-3.6-flash) 和 Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite) 模型现已全面上市,可供开发者在生产环境中广泛使用。特别是Gemini 3.6 Flash,在处理复杂任务方面展现出显著的性能提升,旨在为用户提供更强大的AI能力,并优化模型在多场景下的响应速度与效率。此次发布标志着谷歌在推动其AI模型实用化和普及化方面的进一步努力。
OpenAI Codex项目发布了0.145.0版本,此次更新引入了实验性分页线程历史功能,旨在提升用户在长会话中的高效恢复、搜索和记忆管理能力。新版本还支持子代理功能和持久化名称,进一步增强了代理的灵活性。此外,Codex扩展了导入功能,用户可以轻松迁移Cursor和Claude Code的设置、多组件项目服务器、插件、会话及命令,大幅提升了平台的用户数据管理和跨工具兼容性。
GitHub上发布了Kronos,这是一个专为金融市场语言设计的创新基础模型。它致力于理解并处理复杂的金融文本数据,包括各类新闻、财务报告和研究分析报告,旨在解决传统语言模型在金融专业领域存在的不足。通过深度学习金融特有的术语和上下文,Kronos能帮助金融分析师和量化研究人员进行更精确的市场情绪分析、信息提取和趋势预测,提升金融决策的准确性。
最新研究论文发布了ABot-World-0,这是一款动作条件视频世界模型,其独特之处在于能够在单台桌面GPU上实现实时、长周期闭环的无限互动世界展开。该模型通过整合来自AAA级游戏、模拟引擎及互联网视频的多元数据,学习并生成高度可控的世界动力学。ABot-World-0的推出,为开发下一代交互式AI应用和仿真环境奠定了坚实基础,展示了在有限硬件资源下实现复杂AI仿真的潜力。
Applied Intuition本周正式发布了其代理平台“Dana”,专注于物理AI应用的开发。该平台旨在为工程师提供一套完整的工具,用以构建自主系统,从而加速实体AI技术的落地与规模化部署。Dana平台的推出,有望解决物理世界中AI部署的复杂性挑战,推动智能机器在各类实际场景中的普及,为自动驾驶、机器人等领域的发展提供关键支持。
OmniRoute项目提供了一个免费的AI网关服务,通过单一API端点集成了超过231个AI提供商,其中包含50余个免费模型,极大地简化了多模型调用的复杂性。该平台支持将各类编码助手(如Claude Code、Codex)连接至免费的Claude、GPT、Gemini等主流大模型,显著降低开发成本。同时,OmniRoute采用RTK+Caveman堆叠压缩技术,可节省15-95%的数据传输开销,为开发者提供了统一管理AI服务、优化性能的有效方案。
Voicebox是一个新发布的开源AI语音工作室,为用户带来了强大的语音克隆、文本转语音(TTS)及内容创作功能。该项目旨在民主化高质量语音合成技术,让用户能轻松克隆特定人声,并将文本转化为逼真且情感丰富的语音,进而生成全新的音频内容。Voicebox的出现,降低了内容创作者、开发者以及对个性化语音应用有需求的用户进入门槛,可广泛应用于播客、有声读物和虚拟助手等场景。
GitHub开源项目`outlines`专注于解决大型语言模型(LLM)自由文本生成难以解析和利用的痛点。该项目通过提供结构化输出能力,确保LLM的生成内容严格遵循预定义的格式,如JSON、XML或特定语法结构。这一功能极大地简化了LLM在自动化工作流、数据抽取或API集成中的应用,显著提升了数据处理的准确性和效率,特别适用于LLM应用开发者和数据工程师。
美国财政部长斯科特·贝森特发出警告,指出如果中国AI公司月之暗面(Moonshot)被证实通过「提炼」(distillation)Anthropic的Fable模型来开发Kimi K3,美国政府可能会对其实施制裁。白宫官员早前提出了相关指控,引发了关于知识产权保护和国家安全的新一轮国际争议,显示出AI技术发展在全球地缘政治和经济竞争中的敏感性日益增强。
Uber联合创始人特拉维斯·卡兰尼克(Travis Kalanick)的机器人公司Atoms在新一轮融资中成功筹集了高达17亿美元的资金,由知名风投a16z领投。Uber也参与了此次投资。Atoms公司声称正利用工业AI技术推动全球现代化进程,这笔巨额融资不仅凸显了市场对机器人和实体AI领域的巨大信心,也预示着该领域未来可能迎来快速发展和更广泛的应用。
Anthropic公司在五月份实现了惊人的470亿美元年化收入,远超其2025年90亿美元的预期。Menlo Ventures合伙人马特·墨菲(Matt Murphy)对此表示,这种增长速度在他25年的投资生涯中前所未见,甚至超越了互联网、移动和早期云计算等科技浪潮的初期表现。这一数据凸显了大型语言模型市场巨大的商业潜力和快速增长态势,进一步巩固了Anthropic在AI领域的领先地位。
Anthropic与图书作者达成了一项高达15亿美元的和解协议,这标志着集体诉讼史上最大的版权和解金。该款项主要用于解决因从盗版数据库下载约482,460部作品而引起的争议,而非因AI训练本身。此前,法官Alsup曾裁定,在合法获取的作品上进行AI训练属于合理使用。此次和解虽然金额巨大,但它也间接为AI实验室设定了使用非合法来源数据的风险边界,推动行业对数据来源合规性的重视。
Anthropic团队分享了他们如何巧妙地利用Claude Code来显著加速内部大规模代码迁移项目。此前耗时多年的生产代码库移植工作,在Claude Code的辅助下,现在可以由单个开发者在一个月内完成多达10个代码包的迁移。这一案例不仅展示了AI编程工具在提升开发效率方面的巨大潜力,也为其他企业和开发者提供了借助AI优化复杂工程任务的实用工作流和成功经验。
YouTube用户Riley Brown发布了一段视频,详细展示了他如何通过创新性地使用Codex(推测为OpenAI Codex)来将其改造成为一个高效的「业务增长机器」。视频内容可能涵盖了利用Codex的代码生成、自动化脚本编写或其他AI能力,来优化业务流程、开发新产品或提升市场效率的具体策略和实践经验,为希望 leveraging AI 技术加速企业发展的用户提供了实用的操作指南和启发。
一位博主提出了一个发人深省的观点:虽然人工智能技术显著降低了软件开发的门槛,并加速了制作流程,但伴随着进入门槛的下降,软件的整体质量也可能随之滑坡。这一现象引发了业界对AI普及后软件开发未来走向的深层思考。此观点警示开发者和企业,在拥抱AI带来的效率提升时,需警惕潜在的质量风险,并探索如何平衡AI效率与软件质量之间的关系。
今天的AI产品发布呈现出两大亮点:一是AI Agent相关工具和基础设施的持续完善,从提升开发者体验到提供运行环境,Agent生态日益成熟;二是面向特定领域和用户群体的AI应用深度拓展,尤其是在语音创作、知识管理和专业情报分析等方向,AI正变得更垂直、更易用。
Voicebox 是一个开源的 AI 语音工作室,为用户提供强大的语音克隆、文本转语音及内容创作能力。它允许用户轻松克隆特定人声,将文本转化为逼真且情感丰富的语音,并基于此生成全新的音频内容。该项目旨在降低高质量语音合成技术的门槛,赋能内容创作者、开发者以及任何对个性化语音应用有需求的用户,广泛应用于播客制作、有声读物、虚拟助手或多媒体内容配音等场景。
Lattics是一款模仿大脑结构的知识库管理工具,集成了AI写作辅助和深度研究功能。它旨在帮助用户高效组织、连接知识点,并通过AI技术加速内容创作与学术研究过程。不同于传统笔记工具,Lattics强调知识的「连接」与「演化」,用户可以利用其独特的视图和AI能力,深入探索信息、生成洞见。对于需要处理大量信息、进行内容创作或学术研究的专业人士而言,Lattics提供了一个强大的智能辅助平台。
`worldmonitor` 是一个实时全球情报仪表板,利用AI技术聚合新闻、监测地缘政治事件及关键基础设施动态。它提供统一的态势感知界面,旨在帮助分析师、研究人员或企业迅速掌握全球宏观环境变化,提升决策效率。该工具的亮点在于其AI驱动的实时性与整合能力,能够从海量信息中提取关键信号,为用户呈现高度浓缩且相关性强的情报。适用于需要持续追踪国际局势和重大事件的专业人士、政策制定者及企业风险管理团队。
Arkor是一个开发工具,它使得开发者能够使用TypeScript语言,对开源权重的大型语言模型进行微调(Fine-tune)和部署。这为模型定制化和集成提供了便利,降低了开发门槛。其核心优势在于将复杂的模型操作抽象化,让前端或全栈开发者也能轻松驾驭大模型,从而加速AI应用的开发与落地。Arkor适合希望在自有环境中部署和定制开源LLM,并利用TypeScript生态的开发者。
`pi-web` 项目为 `pi coding agent` 提供了一个直观易用的Web用户界面。它旨在简化用户与AI编程代理的交互过程,摆脱命令行操作的复杂性,通过浏览器即可轻松部署、配置和管理agent,并查看其工作输出。此Web UI极大地提升了 `pi coding agent` 的可访问性和用户体验,让更多开发者能够便捷地利用AI编程代理提高工作效率,尤其是那些更倾向于图形界面操作的用户。
Remote OpenClaw是一个为AI编程代理提供基础设施的平台,集成了超过13,000个MCP(多组件项目)服务器,并提供丰富的技能与插件,旨在赋能AI代理的编码能力。它解决了AI Agent在复杂编码任务中对环境、资源和工具的需求,让开发者能够更专注于Agent逻辑的开发,而非基础设施的搭建与维护。该平台适用于希望构建、测试和部署高性能AI编程Agent的团队和个人。
Buzzy是一款AI驱动的创意协导演工具,旨在通过人工智能技术,为用户提供从初步构思到内容制作指导的全方位辅助,扮演智能副导演的角色。它能够分析创意需求,提供灵感、结构建议,甚至协助脚本创作,显著提升内容创作者的工作效率与产出质量。适用于视频制作人、广告创意人员、作家等,帮助他们突破创作瓶颈,加速项目进程。
AgentManager是一款专为使用Claude Code的用户设计的管理工具,旨在解决因等待用户输入而中断AI编程会话的问题。它通过智能化的会话管理机制,确保用户能够及时响应Claude Code的提示,从而提升AI编程或交互式代码会话的流畅性和效率。该工具特别适合那些依赖Claude Code进行日常开发,并希望最大限度减少中断、保持工作流连贯性的开发者和工程师。
AGINE Academy是一款创新的故事驱动型学习游戏,旨在帮助用户通过沉浸式的实践操作,有效学习和掌握大型语言模型Claude的使用与应用技巧。它将复杂的AI概念和编程挑战融入到引人入胜的叙事中,让学习过程不再枯燥,而是充满互动性和乐趣。该产品特别适合对AI感兴趣的初学者、开发者以及希望以非传统方式提升Claude技能的用户,有效降低了学习LLM的门槛。