VIALS: A Benchmark for Visual Interpretation of Artifacts in the Life Sciences· 生命科学视觉伪影解读基准
In professional life sciences workflows, scientists routinely interpret visual artifacts (…
In professional life sciences workflows, scientists routinely interpret visual artifacts (…
Deep-learning models of anatomy can be numerically plausible yet anatomically impossible, …
Users increasingly turn to large language models for emotional support, yet little is know…
We formalize the Steiner Traveling Salesman Problem (Steiner-TSP) on Graphs of Convex Sets…
The European Union (EU) has emerged as a leading regulatory body in the development of sus…
That a prompt's effect is not a property of the prompt is established: prompts optimised f…
Skills play different roles as an agent's policy evolves: they should first provide learna…
Improving the safety of large language models (LLMs) often comes at the expense of utility…
Large language models often rely on Chain-of-Thought (CoT) reasoning to solve complex task…
Question answering (QA) over long, connected documents remains challenging because relevan…
While recent large language models (LLMs) have achieved promising results on individual pa…
Recently, there has been a great deal of research into improving AI methods and their appl…
Predictive political question answering (QA), such as predicting how a political actor wil…
The rise of agentic AI enables LLMs to access diverse user data, raising critical privacy …
Technology change management in large financial institutions depends on risk assessments t…
Attention masks are relation-level controls: they specify which query--source pairs may in…
Root cause analysis aims to identify the mechanisms responsible for anomalies in complex d…
Large Language Models (LLMs) are moving from code completion toward repository-scale agent…
While multimodal large language models (MLLMs) extend model capabilities beyond text, they…
LLM-as-Judge systems can produce multi-dimensional evaluations, such as trustworthiness, r…
Integrating Large Language Models (LLMs) into the Indian judiciary promises access to just…
Pathology foundation models (PFMs) are increasingly used as general-purpose backbones, yet…
Shape-constrained and physics-informed learning reports an accuracy cost of enforcing a pr…
Class-incremental learning is commonly instantiated as a single-model paradigm, where a un…
Quality gate for AI/Codex-generated pull requests: blocks TODO leftovers, leaked secrets, …
让AI Agent 做真正会思考,有判断的深度调研,而不是会摆信息的汇总报告。一个即插即用的 skill 文件。
A native macOS lab for teaching tiny language models to think — build the architecture, tr…
Local-first, AI-centered, evidence-gated job application workflow
1200+ AI models in your GitHub Copilot Chat — free & forever free. VS Code extension power…
Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per use…
Now exiting stealth mode with a $26 million seed round, Keenable has been building a vast …
TechCrunch talks agents, UX, and reporting to Greg Brockman with OpenAI's head of product.
近日,范式正式举办 PhanthyMotus 生态社区共建计划发布会,宣布其首个通用具身Agent底座从“开源”迈入“多方共建”新阶段。
Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient…
SenseNova U1.5 Lite
8月13日,荣威与火山引擎深度合作的AI原生第一车——家越07正式迎来全球首秀。作为一款面向中国家庭打造的新能源大五座SUV,家越07以“15万级顶格大五座”的实力,将大空间、长续…
<img data-src="https://static.leiphone.com/uploads/new/images/20260825/6a8d3972caa1e.jpg" …
<img data-src="https://static.leiphone.com/uploads/new/images/20260825/6a8d38d7018e7.jpg" …
字节、汇川等已入股未来不远机器人最新一轮融资
AI科研能力从工具级转向系统级
下一代Agent天然会走向云端和更大规模的计算资源
当机器人行业的大量叙事还停留在融资轮次与概念发布时,越疆科技已经交出了一份由产线验收单写成的中期答卷。 8月24日,越疆科技发布2026年上半年中期业绩。报告期内,公司实现营业收入…
2026年8月21日,小猿学习机在京举办“中小学课本学习智能体”首发上线活动。
2026年9月7日–8日,WAIC CONNECT MALAYSIA活动,将在吉隆坡盛大开启!
近日,阿里AI应用搭建产品Meoo(秒悟)打通了自然语言做App的全链路。不懂编程的普通用户只需用自然语言描述想法,Meoo即可自动生成页面与功能,支持真机测试、持续修改与一键打包…
8月25日,豆包工作正式发布。作为豆包面向生产力场景推出的全新 Agent 产品与品牌,豆包工作能围绕用户目标自主拆解任务、调用工具、持续推进复杂工作流程。豆包工作还与飞书深度打通…
<img class="rich_pages wxw-img" data-aistatus="1" data-croporisrc="https://mmbiz.qpic.cn/s…
机器人大脑新SOTA
<img class="rich_pages wxw-img" data-aistatus="1" data-croporisrc="https://mmbiz.qpic.cn/s…
FutureX榜单前十占了仨
<img class="rich_pages wxw-img" data-ratio="0.55" data-s="300,640" data-type="jpeg" data-w…
<img data-src="https://static.leiphone.com/uploads/new/images/20260825/6a8cf7fccd76f.jpg" …
近年来,"AI雷达"成为行业热词。围绕它是什么、为何现在出现、将如何改变毫米波雷达产业,业界正在形成越来越多的讨论。本文基于公开研究、产业实践及超级雷达社区的行业讨论,对其中的核心…
虽然明确机器人创业,但依然在职蔚来
The AI hedge fund went from "the talk of Wall Street" to "subject of federal subpoenas" fa…
OpenAI banned Russia-origin accounts using AI to promote a fake Israel-based think tank an…
Early testers are raving about what Instinct can do, but some say the AI assistant’s sweep…
General Intuition, the startup building a foundation model that trains generalized AI agen…
To generate intelligence at scale, AI factories run continuously, and their economics are …
The next era of AI inference won’t be defined by a single breakthrough chip, network or sy…
According to OpenRouter data, agentic AI workloads consume 15x more tokens than a simple c…
8月24日晚,上海燧原科技股份有限公司(简称“燧原科技”)披露招股意向书、发行安排及初步询价公告等,正式启动发行工作。燧原科技股票代码为“688801”,拟于科创板上市。 燧原科技…
GPT‑5.6 is now available in Kiro, helping developers plan, build, review, and test softwar…
Michael Polansky — better known publicly as Lady Gaga's partner and a former top deputy to…

YouTube · OpenAI

YouTube · Two Minute Papers