跳到正文
今天10月7日周三108 条
  1. Simon Willison

    EmbeddingGemma 2

    My comment on EmbeddingGemma 2 — Hacker News. I really appreciate that EmbeddingGemma 2 is under the Apache 2.0 license. For embedding models in particular, I don't think it makes sense to use a closed, proprietary, hosted-only model. Most applications of embedding models involve calculating thousands or even millions of embedding vectors and storing them for later comparison. If your model is proprietary, the vendor is likely someday going to decide to stop offering that model. They'll have a better model to replace it, but you still need to pay to re-calculate those millions of stored existing vectors. (In April 2024 OpenAI offered to "cover the financial cost of users re-embedding content with these new models" - https://openai.com/index/gpt-4-api-general-availability/ - but I don't think that's something we can rely on from every provider.) Notably, I don't want to host the model myself . I'd much rather pay a provider for a hosted model while knowing that if they ever stop hosting it I can run the open weights version myself - or find another vendor who can do that for me. Tags: google , ai , generative-ai , embeddings , gemma

  2. Simon Willison

    Introducing Mistral Large 4: Le chonk

    Introducing Mistral Large 4: Le chonk Mistral are back in the game. Today they're releasing a preview of Mistral Large 4, a 1 trillion parameter, 49 billion active parameter model trained on their own cluster of 3,800 NVIDIA Grace Blackwell GPUs. The preview is available via their API. They promise to release the open weights model at the "end of this month". The model only supports two reasoning levels - "none" and "high" - via the Mistral API. Here are both pelicans - the "high" one looks better, though surprisingly it only used 2,717 output tokens compared to "none" which used 3,275: On Artificial Analysis it scores 38 , just behind DeepSeek 4.1 Flash, which is a 552B model. It's a huge improvement on last December's Mistral Large 3, which drew this terrible pelican and scored 9 on AA . It's certainly not a Fable-class model, but it's great to see Mistral put out a model that's back to being maybe about 6 months behind the frontier. Via Hacker News Tags: ai , generative-ai , llms , mistral , pelican-riding-a-bicycle , llm-release

  3. The Decoder

    Google claims EmbeddingGemma 2 outperforms rival embedding models twice its size

    Google released EmbeddingGemma 2, an open model with 740 million parameters that converts text, images, video, audio, and code into vectors. It runs on-device, needs only about 191 MB of RAM, and outperforms some competing models twice its size, according to Google. Paired with a small open model like Gemma 4, it can run offline RAG apps without sending data to external servers. The article Google claims EmbeddingGemma 2 outperforms rival embedding models twice its size appeared first on The Decoder .

  4. The Decoder

    Google's new image model Nano Banana 2.1 generates better images for less money

    Google's new Nano Banana 2.1 image model uses Gemini 3.6 Flash and beats the previous Pro model in some benchmarks at a lower cost. But its predecessor also scored well in tests, while Pro often produced better images in practice. The article Google's new image model Nano Banana 2.1 generates better images for less money appeared first on The Decoder .

  5. AWS Machine Learning Blog

    Building a context-aware AI assistant on AgentCore and OpenClaw

    Off-the-shelf AI assistants forget you between conversations. This post shows how to build a personal assistant that accumulates context using OpenClaw on Amazon Bedrock AgentCore runtime, with AgentCore memory turning disposable chats into durable, structured knowledge you can retrieve with metadata filters.

  6. Simon Willison

    Using Parseable with Datasette for OpenTelemetry traces

    TIL: Using Parseable with Datasette for OpenTelemetry traces I saw Parseable in a Show HN today - it's a new observability platform with both an open source (AGPL) Rust implementation (a single ~180MB binary), an "Enterprise" version with extra features and a cloud hosted option. Since Datasette 1.0a41 added OpenTelemetry support (thanks, Alex Garcia), I decided to fire up Codex and have it figure out how to run Parseable and feed it traces from Datasette. Here's my (human-written) TIL showing the patterns that worked, and here's a screenshot of a Datasette trace displayed within the Parseable localhost web application: Tags: datasette , observability , alex-garcia , opentelemetry

  7. Claude Code:GitHub Releases

    v2.1.292

  8. The Decoder

    Wikimedia confirms OpenAI's rogue AI agents edited wikis, tried to compromise tools, and hammered its infrastructure

    According to the Wikimedia Foundation, rogue OpenAI agents edited wikis without permission, tried to abuse a citation tool as a proxy, and may have caused a partial Wikidata Query Service outage through massive crawling. Wikimedia says AI companies need to take responsibility for their agents instead of pushing the burden onto volunteer editors. The article Wikimedia confirms OpenAI's rogue AI agents edited wikis, tried to compromise tools, and hammered its infrastructure appeared first on The Decoder .

  9. Simon Willison

    Mistral Large 4

    My comment on Mistral Large 4 — Hacker News. wren6991 : The benchmark is saturated. Frontier models are tested with an armadillo in fishnet tights jaywalking on Mars. OK well I couldn't resist this one: llm -m claude-opus-5.5 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars' llm -m gpt-6.1-sol 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars' llm -m gemini-3.8-flash 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars' llm -m mistral/mistral-large-4 'Generate an SVG of an armadillo in fishnet tights jaywalking on Mars' Default reasoning levels for each: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

  10. The Decoder

    Microsoft publishes Nobel economist's bearish AI forecast of just 1.5% GDP growth over a decade

    Microsoft published a bearish AI outlook from Nobel economist Daron Acemoglu. He predicts about 1.5 percent GDP growth over ten years and at most five percent of jobs replaced. Bigger models won't move the needle, he argues. What's missing are practical apps that change how work gets done. The article Microsoft publishes Nobel economist's bearish AI forecast of just 1.5% GDP growth over a decade appeared first on The Decoder .

  11. MarkTechPost

    Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model

    Mistral AI has released Mistral Large 4, nicknamed Le Chonk, as a public preview. It is a 1.05 trillion parameter Mixture of Experts model with 49 billion active parameters, native image input, and a 1 million token context window, trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's own European datacenters. The API is live now; open weights ship end of October 2026. The post Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model appeared first on MarkTechPost .

  12. The Verge · AI

    We can’t just change the definition of ‘recording’

    With AI hardware, tech companies are pushing the definition of what does and doesn't constitute a recording. For most of gadget history, it'd be reasonable to assume that a device with a microphone or camera is either recording you or it isn't; it's either on or off, without much gray area. Microphones capture sound. Cameras […]

  13. Pragmatic Engineer

    The state of the tech industry in 2026

    A look into what has changed, what’s still the same – and what’s broken – in the tech industry. The full video and the summary of my keynote at LDX3 New York

  14. Microsoft Research

    What AI gets wrong and what failure teaches us

    Jennifer Neville did not want to go into computer science—but that’s exactly where she landed. Neville discusses the starts and stops that led to her professional sweet spot and her work identifying “surprising failures” making it hard for AI to handle complexity. The post What AI gets wrong and what failure teaches us appeared first on Microsoft Research .

  15. The Decoder

    Insurers brace for millions in claims as AI agents spin out of control

    Insurers are bracing for millions in claims from rogue AI agents, and executives like OpenAI's Sam Altman and Anthropic's Dario Amodei could be personally on the hook for the fallout. The article Insurers brace for millions in claims as AI agents spin out of control appeared first on The Decoder .

10月6日周二
  1. AWS Machine Learning Blog

    Responsible AI governance: How AWS positions customers to align with ISO/IEC 42005:2025

    AWS invests in tools that help customers align with international standards for responsible AI governance. In this post, we explore the AI system impact assessment: what it is, how it improves enterprise-wide risk management, and how ISO/IEC 42005:2025 codifies best practices for conducting and documenting these assessments.

  2. AWS Machine Learning Blog

    Best practices for Amazon SageMaker HyperPod administration and governance

    Learn how to administer Amazon SageMaker HyperPod through Amazon SageMaker Unified Studio while preserving cluster governance. This post shows platform teams how to design infrastructure boundaries, govern access, allocate shared capacity, and operate HyperPod consistently across the organization, project, cluster, and workload control layers.

  3. AWS Machine Learning Blog

    Manage Amazon SageMaker HyperPod Spaces directly from SageMaker Studio

    Data scientists and ML engineers can now create, configure, start, stop, and open Amazon SageMaker Spaces on SageMaker HyperPod EKS clusters directly from SageMaker Studio. Launch JupyterLab and Code Editor environments in a few clicks, without using command-line tools.

  4. Simon Willison

    Scrimshaw Jukebox

    Tool: Scrimshaw Jukebox I wanted to see if Claude Opus 5.5 could compose music, so I tried this : I want you to write some computer game music for me. First design simple text based format for the music and build an artifact that can play it out loud - include some example tracks in that artifact I am looking for music of the quality of the original secret of Monkey Island It leaned a lot harder into the Monkey Island theme than I had intended, but the results are surprisingly good. I wonder if the ability to compose competent music is similar to the 3D graphics thing - a new capability for text models that emerged in the past few months? Would need some careful experiments with other recent and not-so-recent models to confirm if this is new or if they've been able to do this for a while. Tags: ai , generative-ai , llms , claude , vibe-coding

  5. IT之家

    德国交通部长:希望特斯拉 FSD(监督版)辅助驾驶系统能在欧盟获批

    IT之家 10 月 6 日消息,据外媒 Golem 今天报道,德国联邦交通部长施特芬 · 比尔格表示, 他希望特斯拉 FSD(监督版)辅助驾驶系统能在欧盟范围内获批 ,他认为欧盟能够统一审批类似系统具有积极意义。 比尔格表示,特斯拉 FSD(监督版)的审批工作涉及很多有待解决的技术问题, 同时还要划分好责任认定等 。对他而言,推动创新落地是最重要的目的。 这位交通部长还表示,FSD 监督版有望提升德国道路交通安全。他希望这套系统能够尽快在欧洲投入使用。 IT之家注意到,特斯拉曾在今年 4 月明确表示,FSD(监督版)将于未来几个月内在整个欧盟范围内获批。该系统可让驾驶员大幅减少对车辆的控制操作。但实际上,FSD(监督版)仅为 L2 级辅助驾驶系统,驾驶员需要始终做好准备接管方向盘。

  6. IT之家

    《明日方舟:终末地》游戏将支持英伟达 RTX Spark 笔记本,「丹青渡」版本上线

    IT之家 10 月 6 日消息,鹰角网络今日发布《明日方舟:终末地》「丹青渡」版本前瞻视频, 确认本作即将支持搭载英伟达 RTX Spark 处理器的笔记本 。 据介绍,《明日方舟:终末地》将在更新后支持即将上市的英伟达 RTX Spark 笔记本电脑,玩家未来可以使用这些笔记本游玩本作。 IT之家注意到,在 6 月的 2026 台北国际电脑展主题演讲中, 英伟达 CEO 黄仁勋正式宣布推出 RTX Spark PC 处理器 ,首批搭载该处理器的笔记本将在今年秋季推出。 Surface Laptop Ultra :专为全球创客和创意专业人士打造,搭载 15 英寸 mini-LED PixelSense Ultra 微软史上最亮触控屏,以及 Surface 史上最大尺寸触控板,还配有 HDMI、USB-C、USB-A、SD 卡槽和耳机接口。 华硕 ProArt P16 和 ProArt P14:这两款笔记本电脑提供 16 英寸和 14 英寸两种尺寸选择,并有纳米黑和全新的霓虹白两种配色可选。它们配备华硕 Lumina Pro OLED 显示屏,并拥有全天候电池续航能力。 戴尔 XPS 16 Creator Edition:专为创意工作打造,可流畅播放 4K 时间线内容,加快导出速度,并带来更无缝的 AI 工具体验;配备 True Black HDR 600 的 Tandem OLED 显示屏,还内置 SD 卡读卡器和 HDMI 接口。 惠普 OmniBook Ultra 16 和 OmniBook X 14:专为创作者、游戏玩家和 AI 开发人员打造,提供强大的本地 AI 性能和体验,帮助用户加速工作流程。 联想 Yoga Pro 9n:将联想 Yoga 以创作者为中心的特性与英伟达的最新芯片相结合,打造出一款便携、强大且无需插电即可长时间使用的笔记本电脑。 微星 Prestige N16 Flip AI +:融合了轻薄高端的二合一设计、16 英寸 UHD+ Tandem OLED 显示屏、英伟达 AI 加速技术和 99.9Wh 电池。

  7. IT之家

    AMD 股价创历史新高!CEO 苏姿丰称 AI 芯片需求非常旺盛,将持续大幅扩产

    IT之家 10 月 6 日消息,AMD 盘初涨近 3%,报 649.88 美元 (IT之家注:现汇率约合 4,364 元人民币) , 创历史新高 ,市值达 1.06 万亿美元 (现汇率约合 7.12 万亿元人民币) 。 开盘后 AMD 股价有所回落,目前涨超 2%,报 646.48 美元 (现汇率约合 4,335 元人民币)。 AMD CEO 苏姿丰今日在受访时表示,2026 年整体运算市场需求极为强劲,AMD 虽已同步扩充产能,但当前需求依然远高于供给。面对供不应求的现况, AMD 正全力拉升供货,预计 2027 年将大幅扩增产能 。 上月,AMD 股价累计上涨约 30%,总市值首次突破 1 万亿美元 (现汇率约合 6.71 万亿元人民币) 。当地时间 9 月 28 日,AMD 还宣布将以 82 亿美元 (现汇率约合 550.59 亿元人民币) 收购 AI 初创公司 World Labs ,以增强自身技术实力。 相关阅读: 《 AMD 苏姿丰落地中国台湾会见供应链和客户:今早见鸿海刘扬伟,下午再访台积电,私人飞机换成黄仁勋同品牌 》

  8. IT之家

    特斯拉在印度市场交付超一年,仅注册不到 1000 辆汽车

    IT之家 10 月 6 日消息,据 electrek 报道,根据印度电动汽车刊物 ElecTree 汇总的注册数据,特斯拉自 2025 年 9 月在印度开始交付以来, 仅注册了 993 辆汽车 。 上述数据覆盖了 2025 年 9 月至 2026 年 9 月共 13 个自然月;若严格按前 12 个自然月计算,特斯拉交付第一年的新车总注册量仅为 699 辆。鉴于特斯拉官方并不按国家单独披露销售明细,注册量已是最接近真实销量的数据参考。 2025 年 7 月,特斯拉正式进军印度市场,旗下 Model Y 起售价约 70,000 美元 (IT之家注:现汇率约合 47 万元人民币) ,几乎是当时美国本土售价的两倍。 2025 年 9 月是特斯拉在印度的首个交付月,当月新车注册量仅为 69 辆,且在此后的八个月中,其单月表现均未能超越这一水平。 根据 ElecTree 的统计,截至 2025 年底,特斯拉在印度的累计注册量仅为 226 辆;而在 2026 年 1 月至 5 月期间,该数字仅增加了 202 辆,其中 2 月更是跌入谷底,仅登记了 29 辆。 早在今年 1 月,特斯拉便已开始对其首批进口至印度但仍未售出的 Model Y 进行降价促销。到了 5 月, 印度重工业部部长证实特斯拉将不会在印度本土建厂 。这意味着,特斯拉在印度销售的所有车辆都将继续完全依赖整车进口,并需承担相伴而来的高额进口关税。 不过,2026 年 9 月一个月,特斯拉新车注册量达到了 294 辆,创下了特斯拉在印度的单月交付新高。也直接印证了其近期一系列举措的成效 —— 特斯拉全面下调了印度售价,并正式推出了六座版 Model YL。 4 月 22 日,特斯拉在印度正式推出六座版车型 Model YL,起售价为 619.9 万印度卢比 (现汇率约合 43.3 万元人民币) ,并于 6 月启动交付。 这款车身更长、空间更大的六座版车型,其定价反而低于此前在售的五座版 Model Y 长续航版,该车型售价 678.9 万印度卢比 (现汇率约合 47.4 万元人民币) ;目前该版本已被移出在印度销售的产品线。 5 月下旬,特斯拉再次对基础款 Model Y 进行价格调整,直接降价 90 万印度卢比 (现汇率约合 62,888 元人民币) ,售价由 598.9 万印度卢比 (现汇率约合 41.8 万元人民币) 降至 508.9 万印度卢比 (现汇率约合 35.6 万元人民币), 调幅达 15%。