Wednesday, October 7, 2026
25 stories · 241 sources33호
release · Mistral · 6건Korean coverage · GeekNews

Mistral Large 4 preview: 1T-parameter model, open weights due

Introducing Mistral Large 4

Mistral released a preview of Mistral Large 4 via its API, a 1 trillion parameter model with 49 billion active parameters trained on 3,800 NVIDIA Grace Blackwell GPUs, with open weights promised for the end of the month. It scores 38 on Artificial Analysis, up from 9 for Mistral Large 3, just behind DeepSeek 4.1 Flash, a 552B model.

The API offers only two reasoning levels, "none" and "high", which limits how finely you can tune cost and latency, while the coming open weights would make self-hosting an option to evaluate.

Today's 72–7

  1. release · Google DeepMind · 2건

    Google DeepMind releases open multimodal EmbeddingGemma 2

    EmbeddingGemma 2: an open, lightweight multimodal embedding model

    Google DeepMind introduced EmbeddingGemma 2, an open, lightweight multimodal embedding model. Simon Willison praised its Apache 2.0 license, arguing that closed, hosted-only models make little sense for embeddings.

    If you store thousands or millions of vectors for search or RAG, Apache 2.0 weights remove the risk of paying to re-embed everything when a vendor retires a proprietary model.

  2. product · The Verge AI

    Google to limit free Gemini users to Flash Lite from October 9

    Google is about to remove free access to Gemini Flash and Pro

    Starting October 9th, Google will limit free-plan Gemini users to the Flash Lite model, with the standard Flash requiring the $4.99/month Google AI Plus subscription. AI Plus is also losing Gemini Pro, leaving Pro and the "Deep Think" reasoning option to Google AI Pro or Ultra subscribers.

    Developers who lean on Gemini's free tier for prototypes or personal tooling should expect Flash Lite-level output after October 9 and weigh the cost of paying for a subscription versus moving to another model.

  3. research · Hacker News

    Opus 5.5 agents find two room-temperature magnetic semiconductor candidates

    Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates

    A Hacker News post reports that Opus 5.5 agents discovered two candidate materials for room-temperature magnetic semiconductors. The materials are described as candidates, not confirmed discoveries.

    Developers who give agents long-running science tasks should separate candidate generation from experimental validation and budget for the validation cost.

  4. research · Hacker News

    Dust: Pretraining Transformers Without Backpropagation

    Dust: Pretraining Transformers Without Backpropagation

    A post titled "Dust: Pretraining Transformers Without Backpropagation" appeared on Hacker News. Its title indicates an approach to pretraining Transformers that does not use backpropagation.

    If pretraining without backpropagation holds up, it could change activation-memory needs and training pipeline design, so anyone building training infrastructure should check the method's scale and benchmark results first.

  5. product · Anthropic

    Anthropic Expands Its Cyber Verification Program

    Expanding the Cyber Verification Program

    Anthropic announced it is expanding its Cyber Verification Program. The headline alone does not spell out the scope or terms of the expansion.

    Developers using Claude for security work should check Anthropic's announcement for changes to eligibility and coverage before settling on tooling.

  6. product · Simon Willison

    Anthropic moves Cowork's VM from local machines to the cloud

    Quoting Felix Rieseberg

    Anthropic's Felix Rieseberg says the old Cowork ran inference in the cloud but executed tool calls in a VM shipped to the user's computer, which cost disk, battery and performance and stopped work when a laptop was closed. The new version runs both inference and the VM in the cloud, gives each session its own sandbox, and has the desktop app handle file access tool calls when the VM needs something on the device.

    For anyone designing agent runtimes, it shows the trade-off between local and cloud sandboxes: battery cost, whether work survives a closed laptop, and how device files get accessed.

Releases today8

  1. gemini-cli v0.63.0비대화형 모드에서 자율 계획 실행이 가능해지고 인증 무한 루프 등 안정성 문제가 다수 수정됐다트래커·GitHub
  2. claude-code v2.1.292claude plugin install에 --marketplace가 추가되고 Agent 도구에서 서브에이전트의 effort를 지정할 수 있다트래커·GitHub
  3. ollama v0.40.0Apple Silicon에서 MLX 런타임이 지원하는 모델 아키텍처는 기본으로 MLX에서 실행된다트래커·GitHub
  4. transformers v5.19.0텍스트·이미지·오디오·비디오를 하나의 벡터 공간에 임베딩하는 멀티모달 모델이 추가되고, 연속 배칭 attention 설정이 바뀐다트래커·GitHub
  5. ollama v0.40.0Apple Silicon에서 MLX 런타임이 지원하는 모델은 기본으로 MLX에서 실행된다트래커·GitHub
  6. claude-code v2.1.291종료 시 세션 마지막 메시지가 유실되던 회귀 버그가 수정됐다트래커·GitHub
  7. ollama v0.40.0Apple Silicon에서 MLX 런타임이 지원하는 모델 아키텍처는 기본으로 MLX에서 실행된다트래커·GitHub
  8. claude-code v2.1.290플러그인 훅이 서버 도구 호출, 서브에이전트 ID, 조직 승인 수준을 읽을 수 있게 확장됐다트래커·GitHub

Also18

  1. OpenAI will watermark ChatGPT outputs by default—but only in the EU
    Ars Technica AI·05:50
  2. How AI decision models could change content moderation
    TechCrunch AI·05:35
  3. AI computing startup Lambda to raise $4B ahead of planned IPO
    TechCrunch AI·05:00
  4. The next hurdle for AI agents: getting websites to let them in
    TechCrunch AI·04:56
  5. OpenAI agents tried to hack Wikipedia tools and flooded it with traffic
    Ars Technica AI·10/6 21:21·댓글
  6. ChatGPT is adding real cartoonists' signatures to fake New Yorker cartoons
    Hacker News·10/6 07:46·173점·댓글
  7. AI is now capable of developing its own inference hardware
    Hacker News·01:23·103점·댓글
  8. Atlassian and OpenAI expand partnership to turn enterprise knowledge into action
    OpenAI·01:00
  9. Unlocking Earth AI’s planetary geospatial foundation models for global public health
    Google Research·00:05
  10. Advancing computer use with Ironclad
    OpenAI·10/6 19:00
  11. Falcon-Emirati: When an LLM Learns the Dialect, the Culture, and the Nuance
    Hugging Face·10/6 15:44
  12. Open and Emergent Problems in Agentic Privacy and Security: A Contextual Angle
    Google Research·10/6 06:08
  13. Using Parseable with Datasette for OpenTelemetry traces
    Simon Willison·04:07
  14. Hark releases an AI personal assistant with a focus on privacy
    TechCrunch AI·03:22
  15. We can’t just change the definition of ‘recording’
    The Verge AI·01:29
  16. Gemini Call for Me might tell your mom you’re running late
    The Verge AI·10/6 08:09
  17. Scrimshaw Jukebox
    Simon Willison·00:17
  18. Amazon Alexa Plus keeps creepily singing ‘lalala’ for minutes on end
    The Verge AI·10/6 19:53