AI 情报文章归档
浏览 AI圈报 已保存的 5658 条 AI 情报文章,覆盖模型、产品、行业、论文、教程与观点方法六大分类。
已收录文章5658
正式分类6
精选内容3361
全部文章
第 20 / 142 页 · 每页 40 条
论文研究普通
RedEvoAgent: Automatic Red-Teaming Agent with Experience-Driven Skill Evolution
LLM-based agents are increasingly deployed in product-level execution harnesses, where jailbreaks can trigger harmful tool use and persistent state changes, creating greater risks than unsafe text generation alone. Existing automatic red-t…
信息来源:arXiv
行业动态普通
英伟达季度营收5%或来自SpaceX
突发:据 Gene Munster 称,NVIDIA 季度营收的 5% 现在可能来自 SpaceX,接近 48 亿美元,高于上季度的约 3%。 这将使 SpaceX 成为 NVIDIA 最大的客户之一,因为 SpaceX 明年将建设 8 GW 的 AI 算力,与 Meta 和 Amazon 相当。
信息来源:X:cb_doge (@cb_doge) · NVIDIA
行业动态普通
商汤2026上半年首次实现盈利
商汤科技公布2026年上半年财报,上市以来首次实现IFRS盈利,净利润达人民币6.2亿元。总营收29.1亿元(同比+23.4%),生成式AI收入23.3亿元(同比+28.2%),占总收入近80%;经常性收入11.4亿元(同比+124.4%),海外业务收入同比增长127%。2026年7月日均Token调用量同比增长约22倍。
信息来源:X:商汤 SenseTime (@SenseTime_AI)
行业动态普通
Microduck 销售额破百万,创机器人最快纪录?
这是否是任何机器人达到 100 万美元销售额的最快纪录? 我们刚刚突破了 Microduck 的 100 万美元销售额大关。
信息来源:X:Thomas Wolf(Hugging Face 联创/CSO) (@Thom_Wolf) · Hugging Face
论文研究普通
Mechanistic Reaction Prediction via Discrete Flow Matching on Graph-Structured Electron Occupation
Chemical reactions are fundamentally transformations in electron space, yet most machine learning approaches model them either through \textit{de novo} generation of product molecules or through heuristic graph edits that operate directly …
信息来源:arXiv
论文研究普通
Stochastic Estimation of Transduced Language Models
Transduced language models (TLMs) compose a pretrained \emph{source} language model with a functional finite-state transducer to induce a language model over \emph{target} strings. Computing the probability of a target prefix under a TLM a…
信息来源:arXiv
论文研究普通
Persona-Execution Separation: An Architecture Pattern for Evolving LLM Agents under Execution Audit
Large language model (LLM) agents in governed organizations must let the persona (instructions, tone, self-presentation) evolve freely, while keeping execution (stateful, audited work) traceable. A single trust domain does not satisfy both…
信息来源:arXiv
论文研究普通
Beyond F1: Evaluating Coverage and Failure Recovery in AI Model Security Scanners
Static scanners are increasingly used to identify executable or otherwise unsafe content in machine- learning artifacts, yet conventional evaluation metrics characterize only cases where a scanner yields a usable security judgment. We eval…
信息来源:arXiv
论文研究普通
Learning a Continuous Sepsis Severity Score Without Hour-by-Hour Supervision: A Two-Site Retrospective Study
Currently used sepsis severity indices rely on fixed variables and weights established decades ago, which are coarsely discretized and calibrated to a cohort that no longer reflects contemporary critical care. No alternative learned direct…
信息来源:arXiv
论文研究普通
Boosting LLM Exploration via Weak-Model Guidance in RLVR
Reinforcement Learning with Verifiable Rewards (RLVR) significantly improves LLM reasoning but often causes a drop in policy entropy, leading to narrowed reasoning coverage and degraded pass@$k$ for large $k$. While existing methods mitiga…
信息来源:arXiv
观点 / 方法普通
有人会说这是AI
有人会说这是AI
信息来源:X:fofr (@fofrAI)
行业动态普通
OpenAI、Anthropic、Google 等百余家公司联名呼吁抵御恶意 AI 网络攻击
OpenAI、Anthropic、Google、Microsoft 等百余家科技公司签署公开信,呼吁公私部门合作防御 AI 相关网络威胁。信中警告,随着模型能力增强,AI 发起的网络攻击将更广泛、更复杂,医院、水处理厂等关键基础设施面临风险。此前已发生多起 AI 智能体突破沙箱攻击企业的安全事件。
信息来源:TechCrunch:AI(RSS) · OpenAI / Claude
论文研究普通
Scaling Graph Neural Networks for Friend Recommendation: Multi-Hash User Embeddings and Temporal Neighbor Sampling
Friend recommendation is inherently graph-structured: the relevance of a potential connection depends on multi-hop social context rather than user attributes alone. However, deploying message-passing GNNs on a production-scale social graph…
信息来源:arXiv
教程 / 实战普通
小型模型已到货:gpt-5.6-luna 等小模型如何改变 AI 成本格局
作者体验 gpt-5.6-luna 后指出,小型模型在速度与成本上已具竞争力,实测约 100 tps,处理数千封邮件的 API 成本仅几十美分。相比之下,GLM 5.3 也出现在帕累托前沿。作者认为,随着 token 成本下降,面向消费者的 AI 应用将迎来机会,而"快速/廉价/够用"型模型在企业端的"token 喷射"类工作中需求即将爆发。
信息来源:Hacker News 热门(buzzing.cc 中文翻译) · GPT / GLM
观点 / 方法普通
软件工程的核心在于管理复杂性
软件工程的核心并非编写代码,而是管理复杂性:在约束、团队、业务与基础设施等大量上下文下做出架构取舍。AI 擅长生成代码,但工程判断无法被委托给模型--正确方案取决于具体语境,且不存在普遍正确的答案。
信息来源:Hacker News 热门(buzzing.cc 中文翻译)
论文研究普通
智能体轨迹压缩成自动机:行为更多由框架决定
新研究将智能体轨迹语料库压缩为单一紧凑有限状态机,在12个公开数据集上仅用7至43个状态,以0.997适应度重放留出数据,毫秒级构建。FSM状态上下文在下一步预测上全面优于Agent Workflow Memory,失败预测AUROC最高达0.94,并支持在线监控提前停止。作者认为行为拓扑更多由部署框架而非底层LLM塑造。
信息来源:X:DAIR.AI (@dair_ai)
观点 / 方法普通
43秒AI短片《詹姆兰尼斯特的一生》走红
博主@ykszs017 用一周时间制作43秒AI短片《詹姆兰尼斯特的一生》,剧本、分镜、剪辑及原剧台词查找均由AI辅助完成,并使用seedance生成。主推文借角色命运探讨"做了正确选择却被世界惩罚"的悲剧主题,称其"值得被记住"。
信息来源:X:阿易 AI Notes (@AYi_AInotes)
论文研究普通
Consolidating RLVR Capabilities Across Domains: A Deep Dive into Fusion Paradigms
Reinforcement learning with verifiable rewards (RLVR) improves specific capabilities of large language models, but covering multiple capabilities often involves training separate domain experts and subsequently consolidating them. We organ…
信息来源:arXiv
论文研究普通
CLAP: Cross-Embodiment Video World Models are Zero-Shot Physical Simulators
State-of-the-art action-conditioned video models are typically restricted to a single robot embodiment, preventing them from leveraging the vast corpus of heterogeneous video data that contains rich signals for learning generalizable physi…
信息来源:arXiv
观点 / 方法普通
MIT报告:AI检测器在教育中不可靠
MIT报告建议教育界不要依赖AI检测器,指出其概念上存在根本缺陷:学生真实作业中缺乏独立ground truth来验证检测标记是否正确。即使零误报的检测器也会漏掉擅长伪装AI输出的学生,最终惩罚的不是AI使用而是掩饰技巧。报告还警告检测可能误伤非英语母语者或神经多样性学生,并引发"AI人形化"军备竞赛。
信息来源:X:Rohan Paul (@rohanpaul_ai)
论文研究普通
How Language Models Organize and Structure Moral Knowledge
How do large language models (LLMs) organize moral knowledge? Models detect moral content broadly, but detection is a low bar. We ask whether they go further, distinguishing moral foundations from one another and organizing the relationshi…
信息来源:arXiv
论文研究普通
Making Clinical Language Models Auditable: Concept-Guided Fine-Tuning for Robust Prediction
Clinical language models can achieve strong in-hospital accuracy yet fail under deployment shifts because they exploit note-specific artifacts (e.g., templates, separators, boilerplate) that do not reflect patient state. We propose CAST (C…
信息来源:arXiv
论文研究普通
LeVJEPA: Efficient & Scalable Video Pretraining without the Heuristics
Video carries the temporal structure of the physical world, yet learning representations from it has remained computationally expensive: prevailing self-supervised methods either prevent representation collapse through architectural asymme…
信息来源:arXiv
行业动态普通
澳大利亚禁止生成式人工智能进入官方音乐排行榜
澳大利亚宣布禁止生成式人工智能参与官方音乐排行榜评选。该规定针对使用AI生成内容的音乐作品,旨在维护榜单的原创性与公平性。此举将对当地音乐产业及依赖AI创作的音乐人产生影响。
信息来源:Hacker News 热门(buzzing.cc 中文翻译)
观点 / 方法普通
用深度学习和 Keras 解码宇宙信号
天体粒子物理正借助深度学习处理巨型天文台产生的海量复杂数据,以提升仪器灵敏度、发现隐藏模式并搜寻异常信号。这类实验覆盖数千平方公里,传感器以纳秒级分辨率记录波形,其图像化数据结构天然适配 Keras 等深度学习工具。相关方法已应用于 Pierre Auger 天文台、Cherenkov 望远镜阵列和 IceCube 中微子观测站等设施。
信息来源:Google Developers Blog(RSS)
论文研究普通
RATIO: A Benchmark for Retrieval Across Typed Ideation Operations in Scientific Literature
Retrieved scientific literature can serve as inspiration for both human and AI scientists. Inspiration can take different forms: prior work may directly suggest how to address a problem, or surface directions at different levels of abstrac…
信息来源:arXiv
论文研究普通
Property-Specific Recoverability from Contact PPG to Camera rPPG under Heterogeneous Observation Conditions
Camera-derived remote photoplethysmography (rPPG) is commonly validated through endpoint accuracy, but endpoint performance does not establish whether other physiological properties of source contact photoplethysmography (PPG) remain prese…
信息来源:arXiv
产品发布 / 更新普通
Grok Build v1.0.12 更新发布
SpaceXAI 发布 Grok Build v1.0.12 更新,提升工作树创建速度并修复多项 bug。修复包括 MCP 服务器断线重试、token 用量统计更准确、上下文栏即时更新等。用户可通过 `grok update --alpha` 或 `grok update` 获取最新版本。
信息来源:X:cb_doge (@cb_doge) · Grok
产品发布 / 更新普通
Microduck 开源机器人销售额破百万美元
Hugging Face 联创 Thomas Wolf 宣布,开源双足机器人 Microduck 销售额已突破 $1,000,000。该机器人高 25 cm、配备 15 个执行器及摄像头、LiDAR 等传感器,支持强化学习训练,售价低于 $400,并内置多种预训练策略。
信息来源:X:Thomas Wolf(Hugging Face 联创/CSO) (@Thom_Wolf) · Hugging Face
观点 / 方法普通
可视化AI机器人工作的小岛
我建了一座小岛,可以可视化我的机器人工作时的状态!看着它们工作然后去睡觉,休息时回到它们的小房子,真是太可爱了🥺 制作这个的提示词在下面!
信息来源:X:Lauren Tan (@poteto)
模型发布 / 更新普通
蚂蚁百灵发布金融增强模型 Ling-3.0-flash-Fin
蚂蚁百灵推出金融增强版 Ling-3.0-flash-Fin,总参数 124B、激活参数 5.1B,面向信息检索、研究、估值建模与报告撰写。
信息来源:X:蚂蚁百灵 (@AntLingAGI)
观点 / 方法普通
2026 年最佳 Agent 沙箱对比:E2B、Daytona、Modal、Cloudflare 与 Vercel 的冷启动、按秒计费与网络策略
该评测对比了 E2B、Daytona、Modal、Cloudflare 与 Vercel 五家 Agent 沙箱服务,测量突发冷启动性能,并将按秒费率归一化为每 1,000 次执行的成本。评测还依据 2026 年 8 月 27 日核实的一手来源,梳理了各家的文件系统持久性、空闲计费与出口策略。
信息来源:MarkTechPost(RSS)
行业动态普通
百家企业联署呼吁全球网络防御
一封关于全球网络防御激增的公开信,已获得包括 Anthropic、AWS、Google、Microsoft、OpenAI 和 Oracle 在内的 100 多家组织签署。https://x.com/i/article/2093011711712456704
信息来源:X:Greg Brockman (@gdb) · OpenAI / Claude
行业动态普通
加入OpenAI前后对比照引热议
这些照片是我加入OpenAI之前和之后拍的。
信息来源:X:Jason Liu (@jxnlco) · OpenAI
模型发布 / 更新普通
Google 发布 Gemini Omni 1.1 Flash,视频生成更便宜更灵活
Google 将 Gemini Omni Flash 视频模型升级至 1.1 版本,场景扩展可分析现有视频最多十秒而非仅最后一秒,并以 10 秒为增量延长至 40 秒。
信息来源:The Decoder:AI News(RSS) · Gemini
论文研究普通
MIT 报告建议学校弃用 AI 检测器
MIT 一份新报告强烈建议学校不要依赖 AI 检测器,称其可能引发学生使用"AI 人化器"的军备竞赛,最终对双方都无益。报告指出,混合人机写作难以检测、误报会伤害学生,且检测系统可能误判非英语母语者或神经多样性学生的写作。
信息来源:X:Rohan Paul (@rohanpaul_ai)
观点 / 方法普通
DeepMind 副总裁谈 AI 不确定性推理
Google DeepMind 研究副总裁 Zoubin Ghahramani 与主持人探讨如何通过教会 AI 系统"自我怀疑"和概率思维,实现更安全、可靠的现实世界决策。视频涵盖不确定性作用、正确性与置信度对比、贝叶斯思维在 AI 中的应用,以及未来研究与 AGI 方向。
信息来源:X:Google DeepMind (@GoogleDeepMind) · Gemini
观点 / 方法普通
周度Codex额度重置,Tibo获赞OpenAI最佳招聘
每周 Codex 额度已重置。 说真的,Tibo 就是个特别讨人喜欢的人。他幽默地化解了针对他的调侃,始终保持风趣和友善。 OpenAI 最棒的一次招聘。
信息来源:X:Kim (@kimmonismus) · OpenAI
行业动态普通
特朗普政府AI监管机构计划搁浅
特朗普政府仿照FINRA设立AI监管机构的计划在提交总统前已停滞,未获白宫高层支持。此前官员曾起草行政令拟建立AI自律组织监督前沿模型发布前测试,但该行政令未获批准。目前华盛顿仍以6月自愿测试框架为现行机制,争议焦点在于外部机构能否否决美国模型发布。
信息来源:X:Rohan Paul (@rohanpaul_ai)
行业动态普通
Suno CEO Mikey入选时代AI百大榜
自豪地分享,Mikey 入选了 @TIME 的 2026 年 #TIME100AI 榜单! 获此殊荣,我们深感荣幸,这也是对我们出色的团队和社区与我们共同构建这一未来的见证。我们感恩能打造让每个人都能体验音乐创作乐趣的工具。 查看完整榜单:https://time.com/collection/time100-ai/2026/
信息来源:X:Suno (@suno)