Friday, September 18, 2026
25 stories · 264 sources14
product · Hacker News

GLM Built Its Own Inference Infrastructure

GLM Built Its Own Inference Infrastructure

GLM has built its own inference infrastructure in-house. Specific specs or performance figures have not been detailed yet.

A model provider owning its inference stack could shift API pricing and latency, which matters when developers choose which model to build on.

Covered by·Hacker News 115점

Today's 72–7

  1. policy · OpenAI

    OpenAI Publishes Framework for Reporting Model Misalignment

    Our framework for reporting model misalignment

    OpenAI shared a framework for tracking, investigating, and disclosing model misalignment. Alongside it, the company released six reports of unexpected or concerning model behavior.

    Understanding how misalignment is tracked and disclosed gives developers a reference point for designing safety monitoring and risk checks in their own systems.

  2. industry · TechCrunch AI · 2건

    Microsoft Exec Called AI Scraping "Largest Theft of Labor"

    Microsoft exec called AI scraping ‘the largest theft of labor in human history,’ new unredacted filings reveal

    Newly unsealed court filings reveal a Microsoft executive privately called OpenAI's data practices "theft," while both companies scraped paywalled New York Times content to build datasets. Internal emails show the companies feared this practice would gut news publishers in a "doom loop."

    This exposes legal and reputational risk around training data provenance that directly affects licensing deals and data pipeline decisions for anyone building on these models.

  3. product · The Verge AI

    Claude Code Relaunches Cloud-Based Multi-Agent Projects

    Claude Code relaunches Projects to manage multiple AI agents in the cloud

    Claude Code has revamped its Projects feature so users can run multiple AI agents together in the cloud, with a shared memory, goals, and library of files and artifacts. A coordinator directs parallel threads, each its own Claude Code cloud session on a separate branch and repo copy, with overlapping work resolved as a merge conflict like any other PR.

    This shapes how developers structure parallel agent workflows and manage merge conflicts within existing PR-based review processes.

  4. industry · Hacker News · 2건Korean coverage · GeekNews

    Google Launches DeepMind Institute for AGI Safety

    The DeepMind Institute

    Researchers at Google and Google DeepMind have launched the DeepMind Institute (DMI) to study and debate the safe development of AGI and its societal impact. The institute will address not only technical development but also jobs, economic policy, human values, institutions, and governance, with plans to expand participation from outside researchers, the humanities, arts, and government.

    This signals AGI safety and governance work becoming a formalized internal research function at a major AI lab, which matters for teams tracking how policy and ethics considerations may shape future model releases and access.

  5. research · Hacker News

    Research Claims to Break the 1.58-bit Barrier for Ternary LLMs

    Breaking the 1.58-bit Barrier for Ternary LLMs

    A Hacker News post highlights research aiming to push ternary LLMs beyond the 1.58-bit limit. No further details or numbers were shared.

    Developers working on model quantization and inference cost reduction should watch this attempt to surpass ternary LLM compression limits.

  6. product · Hacker News

    OpenSpec, a Lightweight AI Spec Framework, Surfaces on Hacker News

    OpenSpec – A lightweight and configurable AI spec framework

    OpenSpec is presented as a lightweight and configurable AI spec framework. It appeared on Hacker News, drawing attention from developers.

    Developers evaluating tools for defining and managing AI specs may want to check out this new lightweight option.

Releases today4

  1. ollama v0.34.2긴 생성 시 메모리 누적 문제가 해결돼 대용량 컨텍스트에서도 안정적으로 동작한다트래커·GitHub
  2. codex rust-v0.155.0요약 준비 중트래커·GitHub
  3. vllm proto-v0.3.0릴리스 노트 없음트래커·GitHub
  4. claude-code v2.1.274MCP 연결 안정성과 세션 복구 관련 다수 오류가 개선됐다트래커·GitHub

Also17

  1. Nvidia announces native GPU programming in Rust
    Hacker News·9/16 20:15·152점·댓글
  2. OpenAI caught its models leaving notes to successors to hide bad behavior
    TechCrunch AI·05:34
  3. HarnessTax: How Much Does the Harness Matter for Coding Agents?
    Hacker News·9/17 07:10·104점·댓글
  4. Introducing Astra for Law
    OpenAI·9/17 09:00
  5. Introducing the Life Sciences Verification Program
    Anthropic·9/17 09:00
  6. Google announces new experimental "CC" AI agent for families
    Ars Technica AI·05:24
  7. LLMs respond differently to harmful prompts when AI watermarking is used
    Ars Technica AI·03:33·댓글
  8. Show HN: Share your AI Setup, Learn from others
    Hacker News·9/17 22:01·113점·댓글
  9. AI Safety Is Mostly a Sex Cult
    Hacker News·9/17 17:36·110점·댓글
  10. I Don't Like LLMs
    Hacker News·9/17 22:57·105점·댓글
  11. The future of practice: Enabling teachers to create learning interactives with generative UI
    Google Research·05:45
  12. Making global data easier to explore
    Google AI·05:00
  13. datasette 1.0a40
    Simon Willison·9/17 08:51
  14. The fix for rogue AI agents could be more AI
    TechCrunch AI·05:34
  15. Is the AI safety debate about safety or control?
    TechCrunch AI·05:19
  16. UN turns to Google to make its global data ready for AI agents
    TechCrunch AI·05:00
  17. The AI Superintelligence Slowdown
    The Verge AI·04:28

Korea1

  1. 플로우, 국산 협업툴 최초 챗GPT·클로드 앱 동시 입점
    AI타임스·01:26