DawnSift
구독하기
일 · 테크 데일리 · 제49호

2026-08-30

— Tencent open-sourced a 770B-parameter giant model, OpenAI broke with Cursor, and the AI world is full of tension today.

오늘의 TL;DR

Tencent released and open-sourced Hy4 Preview with 770B total parameters, 49B active parameters, and 1M token context; the community has compressed it to 200GB GGUF while retaining about 98% performance. OpenAI announced it will cut off Cursor's access to its models on November 12, 2026, citing distrust that SpaceXAI—after acquiring Cursor—will comply with its terms of service. Sony Music and Warner Chappell jointly sued Anthropic, alleging大规模 copyright infringement in training Claude, with potential damages reaching billions of dollars. vLLM v0.28.0 was released, focusing on optimizing inference performance for Kimi-K3 and DeepSeek V4.

헤드라인

1

Tencent Open-Sources Hy4 Preview: 770B Parameters, 1M Context, Community Compresses to 200GB다중 소스 ×3

Tencent released and open-sourced Hy4 Preview with 770B total parameters, 49B active parameters, a context window exceeding 1M tokens, and weights totaling 1.56TB on Hugging Face. The community has compressed it to roughly 200GB in GGUF format, claiming about 98% performance retention. Why it matters: This is one of the largest open-source models by parameter count, and the 1M context with MoE architecture offers direct value for productivity tasks like long-document processing and code generation; community quantization significantly lowers the barrier to local deployment.

Comments generally acknowledge Hy4's performance and value for money but question its inference speed, open-source definition, and official chart presentation; some also believe its coding capabilities are limited.

2

OpenAI Announces Cutting Off Cursor Model Access on November 12, SpaceXAI Acquisition as Trigger

OpenAI announced it will stop providing model access to Cursor on November 12, 2026, because Cursor was acquired by SpaceXAI. OpenAI stated that based on experience with contract violations by Elon Musk's companies, it cannot be confident SpaceX will comply with its terms of service. Why it matters: Cursor is one of the most popular AI coding tools today, and the model cutoff will directly impact many developers' daily workflows, marking a shift in AI infrastructure competition from technical to commercial and political dimensions.

3

Sony Music and Warner Chappell Jointly Sue Anthropic, Seeking Up to Billions in Damages다중 소스 ×3

Sony Music Publishing and Warner Chappell filed a lawsuit against Anthropic in the U.S. District Court for the Northern District of California, alleging it trained Claude using illegally torrented, scraped, and downloaded copyrighted works. They seek up to $150,000 per infringed work, with total damages potentially reaching billions of dollars. Why it matters: This is one of the latest and largest lawsuits over copyright issues in AI model training data, and the outcome could set a precedent for data-use compliance boundaries across the industry.

4

vLLM v0.28.0 Released: Major Inference Performance Optimizations for Kimi-K3 and DeepSeek V4

vLLM v0.28.0 was released with 584 commits from 270 contributors. Kimi-K3 gained Decode Context Parallel support, fused FlashKDA kernels, and adaptive speculative token budgeting (DSpark TTFT improved by about 60%); DeepSeek V4's sparse MLA now supports plain decode, MTP, and DSpark speculative decoding end-to-end. Why it matters: vLLM is one of the most mainstream LLM inference engines in production, and its continuous optimization for cutting-edge MoE models directly determines service throughput and latency for enterprises with limited GPU resources.

5

Anthropic Research: Claude Trains Claude at $4/Hour, Outperforming $150/Hour Human Researchers

Anthropic published research building an automated alignment researcher system (AAR) based on Claude Opus 4.8, allowing Claude to autonomously read papers, propose solutions, generate data, and train models. It found improvements across all 10 categories of AI safety problems, with some tasks outperforming 28 human safety researchers. Why it matters: This demonstrates a viable path for AI self-improvement, with profound implications for AI safety research paradigms and automated ML workflows, and suggests model capability gains may enter an accelerating loop.

매일 아침, 당신을 위한 테크 다이제스트

웹은 전체 그림을, 구독자에게는 당신만의 것을 — 관심사 맞춤 AI 큐레이션, 개인 RSS 통합, 커뮤니티 반응과 함께 매일 아침 배달. 영원히 무료.

58호 발행 · 매일 150개+ 중 읽을 가치 있는 30개로 선별

AI 소식

개발·오픈소스

Samsung showcased Processing-in-Memory at Hot Chips 2026, integrating MAC units inside LPDDR5X chips.

Comments generally acknowledge PIM's potential but question its practical application, energy efficiency, and software adaptation, noting a lack of killer apps; some still see it as the future direction.

Debian voted to allow 'responsible use of generative AI' in development, maintenance, and documentation.

Most comments support Debian allowing responsible generative AI use, believing developers should be accountable for code; some argue 'responsible' is vaguely defined and worry about quality decline.

커뮤니티 화제

Comments argue GDPR is reasonable but poorly enforced, cookie banners are malicious compliance, and big companies benefit while small ones suffer.

Comments generally agree GDPR is reasonable but poorly enforced, cookie banners are malicious compliance, and big companies benefit while small ones suffer; some see it as legitimizing data trading.

GitHub Trending

tt-a1i/archifyHTML★ 99

Agent skill for beautiful, verifiable architecture, workflow, sequence, data-flow, and lifecycle diagrams—self-contained HTML with motion and crisp export.

A spy satellite simulator in your browser, except the data is real. Live open source spatial intelligence on a photorealistic 3D globe.

Prompt as Code | GPT-Image2 工业级提示词引擎与模板库,470+ 个案例逆向工程,20+ 套工业级模板,并提炼出Skills,持续更新中

FreeToken brings datacenter-scale model serving to your desktop. Run massive models locally, fast and efficiently.

더 볼만한 소식(20건 더)

You can beat SOTA Time Series Anomaly Detection methods with a 100 year old algorithm Time Series Anomaly Detection (TSAD) seems to be one of the hottest topics in NeurIPS, SIGKDD, VLDB etc. Many (perhaps most) papers evaluate on Paparrizos’ TSB-AD-M benchmark… However, I tested these benchmark datasets and found that in most cases I could beat the SOTA TSAD methods with a 100-year-old algorithm, simple Statistical Process Control (SPC). In the attached example, SPC gets perfect results. If we c

The other day I was trying out a distillation of DS4 Pro, and it came with MTP. It was slow as hell on my hardware, barely 2-3 t/s, BUT the speed got bumps every once in a while, and I noticed it was in moments like: United States of America First law of thermodynamics The enshittification of the internet Basically, every time a very predictable phrase came up, it was instantly written. A fun thing to see. But it also has me wondering - would MTP work together with n-grams? Since n-grams are Mar

In mid-August, Ramp published spending data collected from 70,000 U.S. companies: Fable 5 ,the most powerful and expensive model in Anthropic’s lineup, accounts for just 11% of what those businesses spend on the company’s tools. The remaining 79% is worth its weight in gold. With the new releases from Qwen and GLM, we are likely close to Opus 4.8, and certainly ahead of Sonnet and the other LLMs shown at the top of the image. The "anti-open-source crusade" therefore comes as no surprise: it is a

Folks! We're just 50 PRs away from more faster inference . Hopefully by end of year. Experts!, please chip in there. List of Open/Ongoing PRs(and also Discussions) related to CPU/RAM/Disk/Hybrid: [Discussion] RFC: MoE expert cache, VRAM caching of hot CPU-resident experts with hybrid hit/miss execution #24528 AVX2: Speed up large batch size prompt processing of IQ models #27402 llama: add Maple 20B-A1B ternary MoE architecture (CPU)- #27000 ggml-cpu: tiled mul_mat for k-quants- #27851 ggml-cpu:

Came across the subreddit /singularity the other days and many commented that they had 20+ years of experience and hadn’t written one single line of code since 2025. I feel living in a parallel universe, because at my work, no matter how we incorporate AI into our workflow (we literally tried every way people recommended), AI rarely produces the exact or the same quality of codes without decent amount of human developers’ intervention. And at the end of the day, the amount of time we spend revie

In this tutorial, we build an ensemble weather forecasting workflow with NVIDIA Earth2Studio. We install the required Earth2Studio components while preserving Colab’s existing CUDA-enabled PyTorch environment, load the FCN prognostic model, and retrieve atmospheric initial conditions from GFS. We then implement a custom wind-power diagnostic that converts 10-meter wind components into turbine capacity factors, along […] The post Building Custom Batched Ensemble Weather Forecasting with NVIDIA Ea

Pollen Robotics, the Bordeaux robotics team at Hugging Face, opened pre-orders for Microduck — a 25 cm bipedal robot where every movement is a neural policy trained in MuJoCo and exported to ONNX. At $399, it puts the full sim-to-real loop on a desk: 15 motors, camera, LiDAR, two IMUs, and an Apache-2.0 training stack you can retrain yourself. The post Hugging Face Unveils Microduck: A $399 Open-Source 25 cm Biped You Train with Reinforcement Learning appeared first on MarkTechPost .

I've spent the last week or two pondering whether I am in the wrong, or the AI tools are really that uncontrollable, or maybe I am just losing my mind over it. Perhaps someone has similar experiences or found a way to actually do something about it. Thing is, I cannot keep up with all of those new "the" AI tools you are supposed to use to succeed that pop up every other week. Normally I just stick to the simplest things like a CLI agent for my daily work, because I didn't like the UX of AI IDE p

매일 아침, 당신을 위한 테크 다이제스트