jihe520/MathModelAgent
AIMathModelAgent 是专为数学建模设计的 Agent,可自动完成建模并生成可直接提交的论文
AIMathModelAgent 是专为数学建模设计的 Agent,可自动完成建模并生成可直接提交的论文
AIawesome-llm-apps 收录 100 多个免费开源的 AI Agent、Agent Skills 与 RAG 应用
AIClaude-Red 是面向 Claude skills 系统的攻击性安全技能库,每个技能为结构化 SKILL.md 文件
AIFreeCORE 是 TrueNAS 衍生分支,在 FreeBSD 上维护深度整合的虚拟化、Jails 容器与 OpenZFS
AIYuE2 支持符号规划、零样本翻唱与智能体音乐编辑的前沿音乐生成
AIi-have-adhd 是一个让编码 Agent 不再埋没答案、输出更符合 ADHD 阅读习惯的技能
AIspec-kit 是帮助上手规范驱动开发的工具包,主张先定义要构建什么再动手
AIPerplexity 用 GPT-6 Astra 撰写沟通内容、修改软件并监控生产系统,检查频率远低于早期模型
AIbook-to-skill 可将任意技术书籍 PDF 转为 Claude Code 技能,便于学习、查阅和工作使用
AIno-ai-slop 可移除文章中 20 多种 AI 套话模式
AIScrapling 自适应网页抓取框架,支持单次请求到大规模爬取
AIWaymo 车辆监测到乘客持枪,协助旧金山警方逮捕两名未成年人
AIOmniVoice 高质量零样本语音克隆 TTS,支持 600 多种语言
AIGemini API、SDK 及模型智能体交互技能库
AIhyperresearch 智能体驱动研究知识库,自动收集整理网络研究
AIGPT-6 Astra 在 Vending-Bench 测试中收入达 Claude Fable 5.1 三倍
AIomlx 面向 Apple Silicon 的 LLM 推理服务器,支持连续批处理与 SSD 缓存
AI智谱完成约 50 亿美元融资,用于下一代 GLM 基础模型研发
AI华为 Mate XT 2 首发麒麟 9050 Pro,游戏体验比肩骁龙 8 Elite
AIllama.cpp 提交 b10941,缩减 FA 测试规模
AI邬江兴称VLEO超低轨是地球空间最后战略新资源层,2026-2030年是中国换道发展关键窗口期
AI两名英伟达机器人研发员工离职创业,黄仁勋坦诚建言并投资
AI英伟达否认循环融资,称每投资1美元可收回100美元
AIInternLM发布Intern-S2-397B多模态基础模型,面向科学智能与长周期智能体
AI大众安徽与众08猎影版SUV开启交付,限时权益价21.99万元起
AI硅谷从聊天机器人转向资源密集型智能体AI,推动数据中心建设
AIFigma利用AI代理提升安全性
AIQCon上海分享探测式评测管线,重塑大模型评价体系
AIRevolut确认因伪造政府请求导致客户数据泄露
AI本地LLM社区氛围让人重温互联网黄金时代
IT之家 9 月 13 日消息,中国科学技术大学郭光灿院士团队在量子信息安全研究方面取得重要进展。该团队韩正甫、陈巍、银振强、王双等与广东工业大学秦玉文-付松年团队合作,结合量子力学基本原理与相对论时空约束, 实现了具有可证明安全性的量子安全定位 ,可对目标位置进行可信验证,为保障现代信息体系中的位置信息安全提供了新的技术途径。 位置信息是支撑人类活动的基本要素之一。对目标的真实位置进行可靠验证,使定位系统不仅能够确定目标“在哪里”,还
IT之家 9 月 13 日消息,9 月 11 日,有消息称,阿维塔科技董事长王辉已被调去负责长安汽车集团海外业务。对此,阿维塔科技方面向《每日经济新闻》记者回应称:“ 王辉仍为阿维塔科技董事长、长安汽车执行副总裁,分管一部分海外业务。 ” IT之家注意到,据雷峰网 9 月 11 日报道,阿维塔科技正在酝酿新一轮高层调整,现任阿维塔科技董事长王辉已被调去负责长安汽车集团海外业务。在内部系统中,王辉目前仍是阿维塔科技董事长,但他的精力已转向
IT之家 9 月 13 日消息,荣耀产品维护与升级 @荣耀小芳哥 于 9 月 11 日发布称,MagicOS11 在省电上加了个小“彩蛋”——「你的专属省电建议」,升级后,系统会结合近期使用习惯, 识别出一些“耗电较高、但又很少使用”的功能 ,给出针对性的优化建议。 他表示:“没看到提示的朋友也别担心,说明你当前设置和近期习惯比较匹配;想主动优化的话,也可以到【设置 > 电池 > 耗电优化】处理。” 据其分享的截图,系统可为用户检查省电
Since we know Qwen 3.8 27B thinks quite long, but gives at least a good one-shot result where you can leave it to do everything on it own, how does it compare to the older series of models for very simple tasks where you
IT之家 9 月 13 日消息,据英国《卫报》报道,普雷斯顿汽车站对面有一处青年中心,英国政府正在这里尝试解决两大国民经济难题:人工智能与青年失业。 英国政府希望借助前者解决后者,而不是让 AI 加剧失业问题。在兰开夏郡这座城市新落成的社区中心 Vault 内,一场“AI 速成训练营”正在一间教室开展。这是英国西北部地区的试点项目,面向 16 至 24 岁青年提供 AI 相关培训。 训练营讲师站在显示屏前,指导围坐在长桌旁、面前摆着笔记
IT之家 9 月 13 日消息,大眼橙推出了 C5 Ultra 高亮版 / 旗舰版投影仪,将于 9 月 24 日开售,售价 2499 元起(首发价 1799 元起): C5 Ultra 高亮版:2499 元,首发价 1799 元 C5 Ultra 旗舰版:2699 元,首发价 1999 元 两款新品均采用 LCD 显示技术,物理原生分辨率 1920×1080,支持原生 120Hz,投射比 1.2:1,搭载沙姆全屏清晰 2.0 技术,宣称
A law professor spent two years testing how an AI ban, unguided AI use, and structured training affect student performance. The group without AI finished last both years. "I was wrong," the researcher writes, who had ass
IT之家 9 月 13 日消息,据彭博社报道,市场观察人士表示,AI 行业高管呼吁放缓技术研发,短期内可能对芯片厂商及供应链板块的股价形成压力,但长期影响大概率有限,因为算力基础设施的投入规模依旧保持强劲。 周一开盘,半导体企业以及其他 AI 相关股票或将首当其冲遭遇首轮抛售。投资者正在评估,若前沿大模型转向更审慎的研发策略,会不会压缩企业盈利空间。不过,芯片、能源与算力的需求持续供不应求,因此本次股价回调很可能只是短期现象。 IT之家
IT之家 9 月 13 日消息,据央视新闻今日报道,7 月下旬,国家质量强国建设协调推进领导小组办公室会同有关成员单位,首次启动质量督察工作。督察组深入浙江、广东、江苏、山东等地明察暗访,揭开了小商品和灯饰等产业背后隐藏的监管盲区。 报道称, 广东中山 作为“中国灯饰之都”,聚集了古镇、横栏、小榄等产业重镇,民用灯饰销量占全国六成以上。然而,督察组深入多个灯配城暗访暗查发现,部分小企业、小门市违规销售“三无”产品、“套证”冒用认证等问题
IT之家 9 月 13 日消息,据韩媒 ETNews 消息,业内 13 日透露,三星电子与三星显示已敲定在 Galaxy S27 Pro 与 Galaxy S27 Ultra 的 OLED 屏幕上采用代号为“M16”的新型发光材料。 报道称,M16 发光材料已率先应用于本月发布的苹果 iPhone 18 Pro 系列以及上个月上市的谷歌 Pixel 11 系列。此外,vivo 计划于本月底发布的 iQOO 16 也将配备三星显示的 M1
Sam Altman, Elon Musk, and Demis Hassabis back Dario Amodei's call to slow down AI development, at least in part. Altman says OpenAI is pushing its IPO to 2027 over safety concerns. The article Altman, Musk, and Hassabis
IT之家 9 月 13 日消息,今年 4 月,华为常务董事、产品投资评审委员会主任、终端 BG 董事长余承东正式发布了 Pura 90 系列手机。其中,标准版采用了和 Pro / Pro Max 版本不同的设计。 IT之家实测发现, Pura 90 标准版已开放支持了 XMAGE 相机水印 ,包含悬浮样式和相框样式等,补齐了此前影像方面的遗憾。 作为参考,华为 Pura 90 标准版搭载海思麒麟 9010S 芯片,后置全新三角标 Dec
IT之家 9 月 13 日消息,《漫威金刚狼》成了失眠组(Insomniac Games)单人超级英雄游戏佳作矩阵里少见的翻车作品。 《漫威金刚狼》是 PlayStation 近年口碑较差的游戏之一,大量评测对它提出尖锐批评。争议最多的地方之一,就是游戏的辅助引导设计,相关吐槽非常多。 IT之家注意到,Skill Up 发布的《漫威金刚狼》评测引发大量关注,核心观点就是批评这款游戏全程过度手把手引导玩家。 评测者指出,游戏设置了大量直白
IT之家 9 月 13 日消息,华为音悦家 App 在华为应用市场 App Gallery 发布了 6.0.0.320 版本尝鲜升级。新版本新增适配如下机型: HUAWEI MatePad Mini(含悦读版) HUAWEI MatePad Air 12 英寸(第三代) HUAWEI MateBook Pro S HUAWEI Mate XT 2 IT之家注:音悦家是华为自研 App,覆盖作曲、录音、编曲、混音四大数字音乐创作全流程核心
Article URL: https://brew.sh/2026/09/13/homebrew-7.0.0/ Comments URL: https://news.ycombinator.com/item?id=49681545 Points: 63 # Comments: 19
IT之家 9 月 13 日消息,凯越机车昨日宣布,650RR 摩托车正式上市, 售价 33,977 元 : 标准版(光刃白、液态金属银):33,977 元 创世版(赛道赤刃蓝、限量 777 台):33,977 元 凯越 650RR 摩托车搭载 全新自研 645cc 四缸发动机 ,最大马力 124Ps、最大扭矩 69N·m、0-100km/h 加速约 3.4 秒,最高设计车速(极速)260km/h。 这款摩托车的车架按 WSBK 赛事标准
IT之家 9 月 13 日消息,华为官方团队 @智能路由与家庭存储 宣布,智慧生活华为家庭存储适配 HarmonyOS 系统, 新增支持整机备份功能 。 华为手机 / 平板: 6.1.0.125 及以上版本 ,路径:设置 > 关于本机(xx 的手机)>HarmonyOS 版本 / 软件版本 家庭存储: 6.1.0.7 及以上版本 ,路径:智慧生活 App > 存储卡片 > 设置 > 系统更新 智慧生活 App: 17.0.4.313 及
IT之家 9 月 13 日消息,内存涨价潮几乎波及了所有消费电子产品,消息称三星正考虑下月在韩国本土上调全部 Galaxy S26 系列机型售价。 韩国财经媒体 Hans Economy 报道称,三星计划自 2026 年 10 月 1 日起,在韩国本土上调 Galaxy S26、Galaxy S26+ 以及 Galaxy S26 Ultra 的售价。根据发给当地运营商的通知,该系列所有机型统一涨价 149,600 韩元 (IT之家注:现
Comments
IT之家 9 月 13 日消息,据人民财讯报道,9 月 12 日,赛力斯集团董事长张兴海谈到问界造豪华车“方法论”。 外观可以模仿、尺寸可以拷贝、零部件可以采购, 但只靠硬件堆砌难以造出豪华车 。 问界的优势是整车原生体系。在赛力斯魔方平台上,华为智驾座舱、电池等技术与整车架构同步仿真、同步标定、同步验证。 谈及与华为、宁德时代的合作时,张兴海认为不是简单的传统主机厂与供应商的采购关系,而是 业务共建、技术共创、商业共赢 。 据IT之家
IT之家 9 月 13 日消息,宝马近日于印度金奈工厂正式投产 i7 纯电旗舰轿车,印度由此成为继德国之后,全球第二个可下线这款旗舰电动车的国家。 在此之前,宝马 i7 仅由德国丁格芬工厂独家生产。伴随 2027 款 7 系推出, i7 实现印度本土化制造 。本次投产的是 2027 款 i7 50 xDrive,官方定价 1,950 万印度卢比 (IT之家注:现汇率约合 137.2 万元人民币) ;高性能版本 i7 M70 xDrive
IT之家 9 月 13 日消息,据中国铁路官方,9 月 11 日 8 时整,55009 次检测列车从玉林北站驶出,开往岑溪东站。标志着南宁至珠海高速铁路玉林北至岑溪东段(以下简称南珠高铁玉岑段)正式开始联调联试。 南珠高铁玉岑段线路自玉林北站引出,向东经玉林市北流市、容县,终至梧州市岑溪市,全长约 85 公里, 设计时速 350 公里 ,新建容县南、岑溪东两座车站,在玉林北站与 2024 年 12 月已经开通运营的南珠高铁南宁至玉林段相
Article URL: https://jetkvm.com/blog/introducing-jetkvm-mini Comments URL: https://news.ycombinator.com/item?id=49681152 Points: 159 # Comments: 75
IT之家 9 月 13 日消息,索尼今年口碑接连受挫,玩家对该品牌的信心持续下滑。从取消第一方游戏实体光盘的决定,到不再继续支持小岛秀夫的《Physint》项目,每一项决策都让公司深陷争议。 玩家梳理了索尼在第九代主机周期里的其他失误后发现,PlayStation 生态面临更严峻的困境。举例来说,PS5 问世近六年,索尼只打造出 4 个全新系列作品;而 PS4 在整个产品生命周期,一共诞生了 21 个原创 IP。 近期 ResetEra
苹果 iPhone 18 Pro 系列手机现已开启预售,目前京东共支持 2 种购买模式: 现货 10:00/20:00 开抢 +12 期免息 如果下单时未享受任何形式的补贴,到货无需激活即可签收。 京东 苹果 iPhone 18 Pro 12 期免息 9999 元起 直达链接 京东 苹果 iPhone 18 Pro Max 12 期免息 10999 元起 直达链接 预定随时下单 +24 期免息 10 月 10 日发货,需 激活 签收!
亮源新创的Physical Al路线清晰了
出身寒微不是耻辱,放弃自己才是。
vulkan: workaround NV queuesubmit driver bug ( #28830 ) There is a driver bug where two queues on the same VkDevice simultaneously submitting can break some internal synchronization. Until it's fixed, add a mutex around
Agent的下一步是关系型生产力
I have a full model, it's ready to train. It's ~9b parameters. 9.4b to be more exact. That includes a 1/2/3 Engram table, Moonshot's AttnRes modeling, and RoPE / NoPE layering at 3:1 as more or less validated by most maj
Comments
Comments
I can get $5k for the 5090 and the Mac is $5499 before tax. The 5090 has a memory bandwidth of 1.8 TB/s while the M5 Ultra is 1.2 TB/s. Is this a sensible upgrade? Primary use is coding.   submitted by   /u/unchi
opencl: apply the noshuffle row-alignment rule to q4_K, q5_K and q8_0, not just q6_K ( #28575 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47132459 macOS/iOS: macOS Apple
  submitted by   /u/rm-rf-rm [link]   [comments]
I gave DS V4.1 Flash an HLE problem with a bash tool + 2 hours. Hour 1: it wrote three MILP solvers. (225,200) Hour 2: it downloaded the HLE dataset from Hugging Face, found the question, read the answer key (225,600), a
Article URL: https://phys.org/news/2026-08-black-hole-caught.html Comments URL: https://news.ycombinator.com/item?id=49679734 Points: 24 # Comments: 12
Article URL: https://www.historytoday.com/archive/feature/succession-crisis-tore-england-apart Comments URL: https://news.ycombinator.com/item?id=49679647 Points: 37 # Comments: 32
Article URL: https://hyperbo.la/w/aligned-to-whom/ Comments URL: https://news.ycombinator.com/item?id=49679643 Points: 68 # Comments: 43
RSI太危险,得管!
Article URL: https://terrytao.wordpress.com/2026/09/12/after-math/ Comments URL: https://news.ycombinator.com/item?id=49679637 Points: 98 # Comments: 80
暴雪宣布了 FPS 版《星际争霸》,游戏仍然处于早期开发阶段,目标发售时间是在 2030 年。暴雪称,新作是一款开放世界、剧情驱动的科幻射击游戏,故事背景设定在《星际争霸 II》事件发生后数十年,是《星际争霸》宇宙中的一款全新作品。RTS 版《星际争霸》于 1998 年发布,2015 年发布了《星际争霸II》三部曲中的第三部《虚空之遗》,时隔 11 年之后宣布的正统续作不再属于 RTS。FPS 版《星际争霸》游戏设定在 Koprulu
Article URL: https://icm.museum/ Comments URL: https://news.ycombinator.com/item?id=49679459 Points: 121 # Comments: 13
Comments
Comments
Comments
It’s always bothered me that after fine-tuning a model for a project, there isn’t a particularly easy way to host it without either running it locally and keeping a GPU on 24/7 or paying for an entire GPU server. There a
Comments
New website to download models in case HF starts censoring or limiting access.   submitted by   /u/Thrumpwart [link]   [comments]
A few people here mentioned interest in a way to test their custom Pi setups, so I figured I’d drop this here: RoastMyHarness The basic idea is a small engine that sets up an environment to run DeepSWE benchmark tasks us
Article URL: https://yoshuabengio.org/en/publication/why-are-ai-agents-lying-cheating-and-coordinating Comments URL: https://news.ycombinator.com/item?id=49678969 Points: 270 # Comments: 339
Comments
chat : improve parsing of complex types in qwen3-coder ( #28742 ) chat : improve schema support in qwen3 parser cont : clean up grammar a bit Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp
common: add LOG_JSON macro to log structured data ( #28586 ) add LOG_JSON macro fit: add demo LOG_JSON Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47112879 macOS/iOS: macOS
Article URL: https://xeiaso.net/notes/2026/everyone-slowdown-but-me/ Comments URL: https://news.ycombinator.com/item?id=49678683 Points: 528 # Comments: 321
common : implement common_schema internal representation for JSON schemas ( #28736 ) common : implement common_schema types common : implement a json schema optimizer common : reduce optimizations common : refactor json-
Here's a neat thing I had ChatGPT Work with GPT-6 Astra (Max) do this morning: I live at . Figure out 5K and 10K running routes from me that loop from my house. Use OSM data. It worked for 27 minutes and produced exactly
Article URL: https://agentsdock.net/ Comments URL: https://news.ycombinator.com/item?id=49678435 Points: 66 # Comments: 29
jinja : support dot property integer literals ( #28817 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47107718 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (ar
While browsing a benchmark list site, I spotted a recently published 33B parameter model which claimed to beat Qwen3.8 27b on the ArtificialAnalysis (AA) intelligence index. I was obviously excited. But then, while looki
cmake: leave the timestamp out of precompiled headers on clang ( #28816 ) Clang stores the modification time of the precompiled header sources inside the header and refuses the header when they differ. A cached header re
  submitted by   /u/pmv143 [link]   [comments]
Article URL: https://twitter.com/ID_AA_Carmack/status/2098443262214230095 Comments URL: https://news.ycombinator.com/item?id=49677577 Points: 141 # Comments: 174
In May, hundreds of malicious and spam packages were uploaded to RubyGems, causing a serious disruption for the host. Now independent researchers have said that a swarm of OpenAI agents were responsible for the attack. N
Article URL: https://lucumr.pocoo.org/2026/9/12/pdoom/ Comments URL: https://news.ycombinator.com/item?id=49677450 Points: 104 # Comments: 74
I have been battling with my Qwen3.8:27b setup on my rtx 5080 16gb. I am using llama.cpp to run a nvfp4 version of qwen3.8:27b llama-b10699-bin-win-cuda-13.3-x64\llama-server.exe -hf williamliao/Qwen3.8-27B-NVFP4-GGUF:NV
OpenAI CEO Sam Altman confirmed that there would be no OpenAI IPO in 2026 during an interview with Fortune. Over the course of 45 minutes, Altman discussed a variety of subjects including the Hugging Face hacking inciden
California Brown Pelican, in San Mateo County, CA, US The Pacifica Pier shut down at the start of June after a crack in the concrete walkway made access to the pier unsafe. It has since been entirely taken over by pelica
As you all know, Qwen3.8 Flash Next on mainline llama.cpp is still in a pretty experimental stage, but a lot of community forks are trying to get it to work better. There's also a closed-source solution called Halogen (
点击查看原文>
Comments
Article URL: https://withspecific.com/benchmarks/real-swe Comments URL: https://news.ycombinator.com/item?id=49676820 Points: 247 # Comments: 137
While OpenAI has filed confidentially for an IPO, the company will not be going public this year, according to CEO Sam Altman.
Reports of the demise of coders may have been exaggerated.   submitted by   /u/SteppenAxolotl [link]   [comments]
Comments
Anthropic's Dario Amodei and OpenAI's Sam Altman seem to agree that it's time to "pace the frontier." What would that actually look like?
Comments
For a while, I must admit, it looked as if software developer roles like mine were done for. How could we fight against tireless robots? But our industry is slowly realizing that making truly cutting-edge software still
slides In 2025, I [the speaker] found and disclosed a bunch of vulnerabilities in GPG, the most used PGP implementation, and held a talk at 39c3 about it. Some of the bugs ended up getting fixed. This talk describes the
Comments
  submitted by   /u/Thrumpwart [link]   [comments]
Article URL: https://arxiv.org/abs/2609.01877 Comments URL: https://news.ycombinator.com/item?id=49674498 Points: 108 # Comments: 161
Article URL: https://www.bbc.com/news/articles/c14dpgm0rg4o Comments URL: https://news.ycombinator.com/item?id=49674395 Points: 54 # Comments: 107
Coxon, bernie and now this First https://x.com/DarioAmodei/status/2098773920774074715 Then https://x.com/elonmusk/status/2098789109980332057 Then https://x.com/sama/status/2098811563415150910 I think fear mongering appro
墨西哥贩毒集团涉足了加密货币挖矿业务。墨西哥警方在 Puebla 州的 Sierra Norte 地区发现了一个用电量远超周边村庄的矿场,查获了 300 个 GPU、80 个中压终端设备以及 8 个卫星天线。虽然就国际商业规模而言,该矿场的规模相当较小,但这已是自去年年初以来该地区发现的第四个加密货币矿场。根据区块链分析公司 Chainalysis 对流向非法钱包地址的交易量进行的分析,全球范围内非法加密货币交易在 2025 年增长一倍
Article URL: https://high5apps.github.io/josm-plugin-website-wizard/ Comments URL: https://news.ycombinator.com/item?id=49674050 Points: 482 # Comments: 125
Anthropic CEO Dario Amodei says the time has come to slow down AI development and will give third-party evaluators like METR access to its models to help ensure its "adherence to safety practices and commitments." In a w
First off, I know that GLM, Qwen, and DeepSeek absolutely dominate in terms of SOTA Open Source models, and that’s what I use in my personal projects and for school, however, I’m also responsible for deploying local AI o
点击查看原文>
Comments
Comments
Waymo 举报了两名携带幽灵枪的年轻乘客。事件发生在 9 月 3 日凌晨 4 点前,地点是旧金山的 Richmond 区。Waymo 发言人称,它在检测到乘客携带枪支之后,停下了无人出租车,通知了执法部门。旧金山警方拘留了两名未成年青少年,一名男孩和一名女孩,搜查汽车后发现了一支已上膛的 AR 风格突击步枪。两名乘客已被送往少年拘留中心。这不是 Waymo 第一次举报乘客,它在今年 7 月曾举报了玩玩具枪的两名青少年乘客。
Article URL: https://www.righto.com/2026/09/8087-microcode-reverse-engineering-fscale.html Comments URL: https://news.ycombinator.com/item?id=49673580 Points: 114 # Comments: 37
Comments
OpenAI 本周早些时候宣布通过动用约 1 万个 AI 智能体进行长达 88 小时的攻坚,于 9 月 5 日发现了一个 Navier-Stokes 方程失效的特例。Navier-Stokes 问题是克雷数学研究所列出的七大千禧年数学问题之一,它为每道题的解决提供了一百万美元奖金。目前七大问题只有庞加莱猜想确认解决,但解决该问题的俄罗斯数学家格里戈里·佩雷尔曼拒绝接受该奖。克雷数学研究所就此公开声明,表示将根据其评奖流程确认成果。根据克
https://archive.ph/kt50V Comments URL: https://news.ycombinator.com/item?id=49673098 Points: 499 # Comments: 353
Anthropic CEO Dario Amodei is calling for a controlled slowdown in AI development. He warns that recursive self-improvement could threaten the entire internet within six to twelve months and proposes embedded auditors at
Before co-founding Kepler, Vinoo Ganesh led Spark at Palantir and built Project Frontline — a pioneering program for Forward Deployed Engineers. He takes us through the best practices of FDEs.
Most model leaderboards assume a server with powerful GPUs to run models that people daily use. However, my smolbenchmark is the other column: models that fit in 8GB, ranked by: decode speed, tokens per joule, and heat,
ui : add cache ( #28802 ) Signed-off-by: Adrien Gallouët angt@huggingface.co Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47065031 macOS/iOS: macOS Apple Silicon (arm64) mac
Article URL: https://lalitm.com/post/buildprof/ Comments URL: https://news.ycombinator.com/item?id=49672842 Points: 139 # Comments: 28
President Donald Trump is weakening environmental regulations in the name of speeding up the construction of AI data centers, raising health risks for Americans, a cadre of former EPA officials said this week in a briefi
Comments
In a new robotics benchmark, GPT-6 Astra shows major gains in spatial understanding. On StationeryBench, the model completed 7 out of 100 tasks with dual-arm robots, while competitor MolmoAct2 couldn't finish a single on
Article URL: https://darioamodei.com/post/we-must-pace-the-frontier Comments URL: https://news.ycombinator.com/item?id=49672510 Points: 662 # Comments: 933
Nvidia is in talks to invest up to $10 billion in Anthropic's planned IPO, Reuters reports. At a target valuation of $2 trillion, it would be the largest IPO in history. Most of that money will likely end up right back a
Article URL: https://pluralistic.net/2026/09/12/god-in-the-box/ Comments URL: https://news.ycombinator.com/item?id=49672281 Points: 70 # Comments: 36
Reasoning steps like calculation, formula retrieval, and deduction are clearly separable in a model's internal states, especially in the middle layers. That matters for AI safety, because models process more than their v
AuK-Flash: Fast 4-Step Speech Generation and Editing arXiv : https://arxiv.org/abs/2609.08936 Full Paper : https://arxiv.org/pdf/2609.08936 GitHub : https://github.com/Tencent-Hunyuan/AuK Project : https://auk-project.gi
Overly long skill descriptions, blanket reading requirements, and rigid approval rules can get in GPT-6 Astra's way, warns OpenAI's Eric Provencher. More capable models need less hand-holding, so developers should tie in
server : allow model downloads at model limit fix issue #26809 ( #28530 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47055500 macOS/iOS: macOS Apple Silicon (arm64) macOS
Comments
Blog Post : Per-tensor layout maps for GGUF quantization Reddit thread : New tensor type layouts for my GGUF uploads EDIT : Model card has updated things such as Graph, table, text, etc.,   submitted by   /u/pmtt
包括陶哲轩、新晋得主邓煜在内的 25 名菲尔茨奖得主发表公开信《A Severe Misalignment of AI in Mathematics》,批评 AI 公司最近的所作所为。公开信称,“过去几个月 LLM 的数学能力有飞跃式提升,甚至达到了能解决数学领域重大悬而未决问题的地步。但各大 AI 公司仅仅将解决数学问题作为基准测试(Benchmark)推动的技术竞赛,却对数学这门科学以及整个数学界构成了伤害。AI 公司的目标与数学界
Hey everyone! I'm curious to hear from people that use a combination of cloud-based frontier models and local ones for development. I'm planning to set something similar up and wanted to hear about actual examples of thi
太初(杭州)集成电路有限公司新一代超智融合计算系统元碁Hypertintellix入选“算力中国·年度卓越成就”。
Article URL: https://tedium.co/2026/09/11/ilands-agents-email-spam-kaixin-tang/ Comments URL: https://news.ycombinator.com/item?id=49671159 Points: 114 # Comments: 55
Unitree might be the world’s most important robotics company.
OpenAI has spent the last few years planting flags across the increasingly difficult terrain in mathematics. This week, it claimed one of its biggest prizes yet: a solution to a legendary Millennium Prize problem. In nor
Plus: The US disrupts the internet’s biggest black market, a Conti ransomware hacker gets prison time, Meta fails to stop AI-generated videos of child abuse.
点击查看原文>
Applied science work, from workflow design, data pipeline, results analysis, article/reports writing and data publishing online. 5 projects I did in the past replicated from start to finish. 3x to 4x more total wall time
点击查看原文>
点击查看原文>
In May 2026, OpenAI agents uploaded more than 2,000 malicious packages to RubyGems, found an unknown security vulnerability on their own, and tried to steal API keys. The apparent goal was pointless: scraping publicly av
点击查看原文>
点击查看原文>
Google Research has released TimesFM-3, a forecasting model that analyzes time series alongside related data and known future events like sales promotions or weather forecasts. Instead of predicting the future step by st
Claude越界攻击真实系统,并非只是测试系统的设置问题,模型本身的安全问题也出了问题。
触觉、记忆、Ego数据、自进化……这个世界模型全都有
I find this new model at HF: "Built for demanding work. A 262 144-token context window, adjustable reasoning effort, tool calling, and text, image and video understanding. Architecture Agnes-3.0-Flash is a hybrid-attenti
FrontierMath Tier 4,饱和了
冲刺港股IPO
We agree with Sebastian: this should have been DeepSeek v5
25位菲尔兹奖得主联名吹哨
Comments
OpenAI agents carried out an undisclosed attack on RubyGems is a new bombshell report from Spencer Kitts, Thomas Larsen, and Sydney Von Arx - three of the four authors of the report on the agent attack on disused wikis (
Article URL: https://twitter.com/venturetwins/status/2098456905526211026 Comments URL: https://news.ycombinator.com/item?id=49667253 Points: 64 # Comments: 71
Comments
"Dedicated launch is pretty essential for us for most of our missions."
The round for the two-year-old startup is coming together months after Mecka announced its Series A.
So you want to use OpenRouter? One of OpenRouter's selling points is that it "handles fallbacks automatically and picks the most cost-effective option for each request", so you can call a single API endpoint for a model
Simple games gain rich strategies in the face of noise.
Article URL: https://www.dwarkesh.com/p/john-beren-charlie Comments URL: https://news.ycombinator.com/item?id=49665711 Points: 117 # Comments: 116
The strain lurking in the inflatable structure was hypervirulent.
点击查看原文>
点击查看原文>
Tan wants smaller, American open-weight AI labs to use the same kind of training techniques on American frontier AI labs, giving the U.S. a more robust set of open-weight options that aren’t Chinese.
Twenty-five leading mathematicians signed an open letter arguing that AI labs are threatening their intellectual work.
New Mexico's Supreme Court is punishing a lawyer for including AI-fabricated witnesses and fake police testimony in an appeal for his client's murder conviction, according to a report from Reuters. In a filing on Wednesd
Only one week left to secure your exhibit table. Tables are limited and can sell out before the September 18 deadline.
The Department of Energy declared an "emergency" when none existed.
The absolute last chance to apply to host an official Side Event during TechCrunch Disrupt 2026 is tonight, September 11, at 11:59 p.m. PT.
Employees at the world's leading AI labs are saying there's a real possibility that advanced AI could destroy humanity. Are they right? Or is this more scaremongering and hype? Join MIT Technology Review executive editor
While K3's usage figures have declined slightly in recent months, OpenRouter data currently shows as many as 300 billion tokens being generated each day by K3 models on the system.
"I didn't know that AI could hallucinate facts," New Mexico defense lawyer says.
点击查看原文>
Comments
The proposed class action alleges Meta illegally harvested people’s Facebook and Instagram photos to train its AI image-generation models and to build its unreleased “NameTag” face recognition feature.
Analysis revealed seed proteins from sesame and moringa, which held religious and symbolic significance.
An Anthropic researcher resigned this week, warning in a post on X that the company is “racing straight to self-improving superintelligence and gambling with our lives”. The company's own alignment le
A health plan CFO closes the month after the usual round of extracts, spreadsheets,...
Multi-agent systems fail in ways traditional monitoring misses. This post presents a dual-layer approach to monitoring production agents: Amazon Bedrock AgentCore Evaluations for continuous quality scoring and AWS DevOps
Comparing models on dollars per million tokens misses what production workloads actually pay for: outcomes. This post shares an open-source benchmarking harness that measures cost per correct answer, agent trajectory cos
Learn how to build and deploy an MCP App with interactive HTML widgets on Amazon Bedrock AgentCore. Because MCP Apps is a host-agnostic standard, the same server delivers the same rich experience across AI hosts like Cha
Renewables pledge won’t change the Oracle and OpenAI data center’s use of gas.
点击查看原文>
This is the first post in a new series exploring how Lakeflow Connect brings fully...
Production code written by Claude should have a higher bar than if it was written by a human. At Anthropic, we have many guardrails in place to make sure this is happening: lots of lint rules, lots of tests, Claude-drive
https://terrytao.wordpress.com/2026/09/11/a-severe-misalignm... https://www.economist.com/science-and-technology/2026/09/11/... , https://unwall.app/www.economist.com/science-and-technology/... Comments URL: https://news