| 来源 | 时间 | 内容(英文原文 / 中文翻译) |
|---|---|---|
| X | 09-07 12:00 | @guzelonur @GoogleAIStudio Any examples you can send me? Will fix 有例子可以发给我吗?我会修。(回应Google AI Studio界面bug反馈) @OfficialLoganK · 查看原文 |
| X | 09-07 11:55 | @alexolegimas I think there is a difference between Software and software. I think everyone is going to build their own "small s" software. I do not think everyone is going to build their own "big S" Software. There is a lot of small s software needs out there 我认为Software和software是有区别的。我觉得每个人都会去构建自己的「小写s」software。但我不认为每个人都会去构建自己的「大写S」Software。外面有大量小写software的需求。 @emollick · 查看原文 |
| X | 09-07 12:41 | AGI, eh? “GPT Astra is a terrible executor. It writes awful code. It makes tons of mistakes. It can't build even the most trivial sequence of actions — it acts first and thinks later. It doesn't consider consequences. Not once has it given me a bug-free result on the first try. AGI?是吗?「GPT Astra是个糟糕的执行者。它写的代码很差,错误一大堆。连最琐碎的动作序列都搭不起来——先动手后思考,不考虑后果。从来没有一次能在第一轮就给出无bug的结果。用它做项目简直痛苦。」 @GaryMarcus · 查看原文 |
| X | 09-07 10:55 | an important essay on the state of AI and the choices and challenges ahead: for openai, the field, and the world.
we're in the AGI era, and need to approach the next few years with seriousness, thoughtful deliberation, and navigating the collective challenges together. we can 一篇关于AI现状以及未来选择与挑战的重要文章:对OpenAI、整个领域,以及整个世界而言。我们已进入AGI时代,需要以严肃态度、深思熟虑来应对未来几年,并共同应对集体挑战。我们完全可以把未来变得比过去好得多。 @gdb · 查看原文 |
| 来源 | 时间 | 内容(英文原文 / 中文翻译) |
|---|---|---|
| Hacker News 应用落地 | 09-07 12:49 | Show HN: Engrim – A universal, local-first SQLite memory engine for AI CLIs Hacker News · 查看原文 |
| Hacker News 监管政策 | 09-07 19:46 | Smartphone makers don't bother to comply with EU repairability requirements Hacker News · 查看原文 |
| Hacker News 安全伦理 | 09-07 12:38 | I refused to train the AI that could replace me Hacker News · 查看原文 |
| 新闻 其他 | 09-05 00:42 | Variational Continuation for Double Pendulum Periodic Orbits We present a Hessian-based approach to numerically continue periodic orbits in dynamical systems. A loop (periodic orbit candidate) is parametrized as a Fourier series; a loss function is defined based on the deviation of the loop from the physical differential equations. Unlike previous work relying on hand-derived Jacobians, our method automates the process by leveraging automatic differentiatio |
| 新闻 应用落地 | 09-05 00:44 | Does Your Agent's Memory Survive a Model Upgrade? A Controlled Study of Memory Portability Model upgrades are routine; memory migrations are not. An agent can keep the same memory store and still forget: a new model may interpret old notes differently, mixed embedding versions may break retrieval, and repair may fail without the original evidence. We compare memory as the same history is preserved verbatim for long-context reading (LC-RAW), divided into chunks for retrieval-augmented ge |
| 新闻 应用落地 | 09-05 00:49 | Who Should Grade My Work? Student Perspectives on Transparent AI-Assisted Writing Assessment in Higher Education The integration of GenAI tools into higher education assessment raises important questions about how students understand, interpret, and respond to AI-mediated evaluation. As instructors increasingly explore AI tools for providing feedback, prior research has examined whether GenAI-generated feedback improves writing performance and how students perceive its usefulness; comparatively little is kno |
| 新闻 应用落地 | 09-05 01:08 | Distill Globally, Adapt Locally: Reasoning Distillation and Product-Type Test-Time Training for Scalable Trade-Up Recommendation Trade-up recommendation identifies higher-quality alternatives that preserve a customer's purchase intent while offering upgraded benefits. Large language models (LLMs) can reason about such distinctions, but applying them directly to hundreds of millions of product pairs is operationally impractical. We introduce a two-level framework that distills LLM reasoning into an efficient non-generative s |
| 新闻 应用落地 | 09-05 01:08 | Design Docs Are All You Need: An AI-native Machine-Learning Performance Tool Machine-learning performance modeling is a uniquely hostile terrain for long-lived software: the assumptions baked into today's abstractions are invalidated by tomorrow's models and systems, forcing perpetual refactoring of performance-modeling frameworks. Meanwhile, AI coding agents have become fast and capable enough that regenerating an entire library is cheaper than paying down the tech debt o |
| 新闻 安全伦理 | 09-05 01:16 | When LLM Decompilers Recompile More and Preserve Less Decompilation recovers high-level source from compiled machine code and serves as a foundation for security tasks such as vulnerability detection and malware analysis. Traditional decompilers like Ghidra and Hex-Rays expose whatever they cannot resolve as visible placeholders and often emit pseudocode that will not compile or execute; LLM-based decompilers produce clean, idiomatic C and are now ju |
| 新闻 应用落地 | 09-05 01:26 | CUA-Universe: A Scalable and Dynamic Environment for Hybrid GUI+CLI Agents Computer-use agents have advanced on benchmarks like OSWorld and AndroidWorld, but still act mostly through the GUI, often producing inefficient trajectories. Real-world computer work is hybrid, combining visual-state inspection with precise, high-throughput command-line operations, so capable agents must coordinate both modalities over shared application state. Yet scalable hybrid environments re |
| 新闻 应用落地 | 09-05 01:27 | What Matters, When? Diagnosing and Improving Conditional Visual Grounding in Visuomotor Imitation Policies Visuomotor imitation policies can achieve high performance under in-distribution visual conditions yet fail when visually similar objects or receptacles are introduced. We study this behavior as a problem of conditional visual grounding: the visual target required for successful control changes with the manipulation phase and, in more complex tasks, with the observed task state. Using Action Chunk |
| 新闻 安全伦理 | 09-05 01:32 | Molecular Déjà Vu: Digit-Level Retrieval of Published Values in Frontier Language Models Large language models (LLMs) are increasingly evaluated on molecular property benchmarks, but accuracy cannot distinguish a model that predicts a property from one that retrieves a published number. We audit 22 frontier models on 12 regression benchmarks for verbatim retrieval and find that it is widespread but relatively benchmark-specific: on five datasets more than $50\%$ of the LLMs show verba |
| 新闻 应用落地 | 09-05 01:33 | Reflection-aware Generative Novel View Synthesis We propose Ref-GeNVS, a training-free, reflection-aware method for generative novel view synthesis (NVS) in mirror scenes. Existing multi-view diffusion models often fail to recognize the mirror in the scene and cannot exploit reflected content for scene generation. To fix this issue without additional training, our key idea is to treat a mirror image as two complementary views. From input images, |
| 新闻 其他 | 09-05 01:37 | Necessary or Sufficient? Evaluating LLM Explanations With Behavioural Evidence LLM decision components that can operate within agent workflows often produce action-relevant recommendations or judgements together with explanations. Operators may use the named factors to monitor a system, diagnose errors, or decide when to escalate an output. Such use assumes that the explanations agree with the component's observable decision behaviour. We test two interpretations of the name |
| 新闻 应用落地 | 09-05 01:44 | Multi-Step Tool-Calling over Korean Open Public APIs: A Benchmark and a Data-Synthesis Recipe Data-sovereignty regulations increasingly require public institutions to deploy open-source, on-premise LLM agents that chain multiple tool-calls across live government APIs. However, open-source models consistently underperform in this multi-step setting, and no existing benchmark measures the gap. We introduce the Korean Open Public API Benchmark (KOPA-Bench), comprising 145 real-world tasks. To |
| 新闻 其他 | 09-05 01:44 | A Deep Generative Model for Synthesizing Labeled Wireless Signals Wireless signals with position-related labels are pivotal for both performance evaluation and model training in the realm of wireless sensing. However, acquiring real-world datasets is often challenged by significant measurement and labeling costs. Traditional methods for synthesizing labeled wireless signals typically rely on environmental models, leading to extensive hyper-parameter tuning and i |
| 新闻 其他 | 09-05 01:47 | Same Trajectory, Contradictory Rewards (ROBORMBENCH): Paraphrase Fragility in Vision Language Reward Models Vision-language models are increasingly used as reward functions for robotic learning, but this role requires paraphrase invariance: the same trajectory should receive the same reward under semantically equivalent goal descriptions. We show that current VLM reward models often violate this property. Paraphrasing the instruction alone can substantially change predicted progress scores, and can even |
| 新闻 应用落地 | 09-05 01:50 | RegionFed: Federated Learning for Personalized Query Understanding in Heterogeneous Retail Environments Retail search systems serve diverse geographic regions with distinct query patterns, vocabularies, and product preferences, creating significant data heterogeneity that challenges both privacy-preserving training and model personalization. Federated learning offers a natural solution for privacy, but standard FL methods produce global models that sacrifice regional performance, while existing pers |
| 新闻 应用落地 | 09-05 01:51 | Diffusion TV: Experiencing Diffusion Models through Tangible, Embodied Interaction Diffusion TV is an interactive AI art installation that offers a tangible and embodied experience of diffusion models through a modified CRT TV. By physically manipulating the TV's antenna, audiences control the clarity of AI-generated images and sounds, metaphorically enacting the denoising process that underlies diffusion-based generation. Using the tuning knob, participants switch between three |
| 新闻 应用落地 | 09-05 01:52 | WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data Recent advances in wearable sensing enable continuous monitoring of physiological and behavioral signals, yet existing benchmarks rarely evaluate whether AI systems can reason over a real user's longitudinal wearable record. We introduce WearableQA, a benchmark comprising 4,084 10-option multiple-choice questions constructed from the wearable time series, blood biomarkers, and demographics of 200 |
| 新闻 应用落地 | 09-05 01:59 | UniMate: One Unified Model to Animate Diverse Skeletons Recent advances in automatic rigging now deliver animation-ready 3D assets at scale, yet generating the motion to drive them remains a bottleneck. Existing learned animators are topology-constrained: they rely on category-specific templates or require per-skeleton fine-tuning and reference motions at inference. We present UniMate, a unified foundation model that synthesizes articulated motion for |
| arstechnica.com 资本动向 | Mon, 07 Sep 2026 | “If there’s no hydrant for me,” Matisz said, “that fire is going to burn under their asses.”
Loading
Listing image for first story in Most Read: Tesla’s Cybercab has been deployed, and it’s already under investigation [...] AI companies whose compute demand the facility exists to serve. [...] energy, directed inquiries there. arstechnica.com · 查看原文 |
| community.openai 应用落地 | Sun, 06 Sep 2026 | | Looking for Feedback: Recombinant AI™ Plugins / Actions builders chatgpt , plugin-development , chatgpt-plugin , feedback , plugins | 50 | 6865 | November 6, 2023 |
| What if AI wasn’t designed to answer us, but to challenge how we think? Community chatgpt | 27 | 1670 | January 22, 2026 |
| How do you handle overlapping Codex skills in larger skill catalogs? Codex codex , skills , codex community.openai.com · 查看原文 |
| anthropic.com 模型发布 | Mon, 07 Sep 2026 | ## Safety, security, and alignment
AI models’ agentic capabilities have become much more powerful over the past two years. But as we’ve documented, greater autonomy comes with new risks. Work on safety, security, and alignment needs to advance at the same pace as AI capabilities. Yesterday, we published a report describing how we are improving our own alignment and security efforts [...] In addit anthropic.com · 查看原文 |
| Hacker News 应用落地 | 09-07 12:18 | Coop – Isolated VM Environments for Running Claude Code and Codex Hacker News · 查看原文 |
| Hacker News 应用落地 | 09-07 09:18 | Ponytail: Lazy Senior Engineer Skill Hacker News · 查看原文 |
| Hacker News 模型发布 | 09-07 13:23 | Nvidia's Jensen Huang says 'AGI has arrived' and congratulates OpenAI Hacker News · 查看原文 |
| Hacker News 算力芯片 | 09-07 17:26 | Speculative Decoding in vLLM on AMD GPUs Hacker News · 查看原文 |
| Adgully.com 模型发布 | 09-07 12:37 | Google DeepMind unveils WeatherNext 3 with real-time hourly AI forecasting - Adgully.com Google DeepMind发布WeatherNext 3,提供实时逐小时AI天气预报 - Adgully.com Adgully.com · 查看原文 |
| The Tech Buzz 资本动向 | 09-07 09:06 | Google DeepMind Backs 16 APAC Climate AI Startups - The Tech Buzz Google DeepMind支持16家亚太气候AI初创企业 - The Tech Buzz The Tech Buzz · 查看原文 |
| TNGlobal 资本动向 | 09-07 16:02 | Google DeepMind picks 16 environmental-AI projects across Asia-Pacific for inaugural accelerator - TNGlobal Google DeepMind遴选亚太16个环境AI项目入选首期加速器 - TNGlobal TNGlobal · 查看原文 |
| blog.google 资本动向 | 09-07 09:08 | Backing 16 green AI projects in Asia-Pacific - blog.google 支持亚太地区16个绿色AI项目 - blog.google blog.google · 查看原文 |
| Fortune 资本动向 | 09-07 18:35 | Exclusive: Ineffable Intelligence adds 6 "cofounders," hiring from Google DeepMind and InstaDeep - Fortune 独家:Ineffable Intelligence新增6名“联合创始人”,从Google DeepMind和InstaDeep挖人 - 《财富》 Fortune · 查看原文 |
| GlobeNewswire 模型发布 | 09-07 21:00 | Cipheras Group - Apex AI Fund Deploys 290-Billion-Parameter Proprietary Large Language Model - GlobeNewswire Cipheras Group - Apex AI基金部署2900亿参数自研大语言模型 - GlobeNewswire GlobeNewswire · 查看原文 |
| Nature 其他 | 09-07 17:37 | Causal evidence that language models use confidence to drive behaviour - Nature 因果证据表明语言模型利用置信度驱动行为 - 《自然》 Nature · 查看原文 |
| Saint-Gobain 应用落地 | 09-07 16:00 | Saint-Gobain appoints Annica Hagen as Chief Artificial Intelligence Officer and accelerates AI deployment - Saint-Gobain 圣戈班任命Annica Hagen为首席人工智能官并加速AI部署 - 圣戈班 Saint-Gobain · 查看原文 |
| Bruegel 监管政策 | 09-07 16:08 | Tough choices ahead as the US pushes a global artificial-intelligence divide - Bruegel 美国推动全球人工智能鸿沟,艰难抉择在即 - Bruegel Bruegel · 查看原文 |
| hawaiitribune-he 应用落地 | 09-07 18:05 | UH launches free AI course aimed at building literacy across Hawaii - hawaiitribune-herald.com 夏威夷大学推出免费AI课程,旨在提升全夏威夷AI素养 - hawaiitribune-herald.com hawaiitribune-herald.com · 查看原文 |
| Gizmodo 应用落地 | 09-07 17:30 | Spare a Thought for the Kenyan Essay Writers Who Used to Help Students Cheat Before AI - Gizmodo 想想那些在AI出现前帮学生代写作弊的肯尼亚作文写手吧 - Gizmodo Gizmodo · 查看原文 |
| thetimes.com 安全伦理 | 09-07 17:10 | ‘Godfather of AI’ warns of catastrophic consequences for humanity - thetimes.com “AI教父”警告人工智能将对人类造成灾难性后果 - thetimes.com thetimes.com · 查看原文 |
| ABC News & Headl 安全伦理 | 09-07 17:58 | French burglar used ChatGPT to locate Canberra home with luxury goods - ABC News & Headlines – Australian Broadcasting Corporation 法国窃贼利用ChatGPT定位堪培拉藏有奢侈品的住宅 - ABC新闻与头条 – 澳大利亚广播公司 ABC News & Headlines – Australian Broadc · 查看原文 |
| Reuters 安全伦理 | 09-07 20:35 | AI could pose 'existential' risk to humanity, UN rights chief warns - Reuters AI可能对人类构成“存在性”风险,联合国人权事务高级专员警告 - 路透社 Reuters · 查看原文 |
| 来源 | 市场 | 榜位 | 时间(北京) | 标题(点击查看原文) |
|---|---|---|---|---|
| 澎湃新闻 | 综合 | 10 | 09-07 21:14 | 讲武谈兵|军用人形机器人走向战场,“终结者”来临? |
| bilibili 热搜 | 综合 | 11 | 09-07 21:14 | 苹果和OpenAI为啥反目成仇 |
| bilibili 热搜 | 综合 | 21 | 09-07 21:14 | AI做3D的正确打开方式 |
| 微博 | 综合 | 24 | 09-07 21:14 | 华为新一代鸿蒙AI智能手表发布 |
| 百度热搜 | 综合 | 24 | 09-07 21:14 | 军用人形机器人走向战场 |
| 抖音 | 综合 | 25 | 09-07 21:14 | 姜月初AI拼豆手速比斩妖还快 |