(오늘의 짤방: RTX 3090 owners tonight will be running Kimi_K3_3T_Q_0.001_K GGUF via @TheAhmadOsman)
- 빅데이터/인공지능
- Anthropic Opus 5 출시 (anthropic.com)
- Claude Cookbook: 에이전트부터 RAG·멀티모달·운영까지 (platform.claude.com)
- Claude 5 모델을 위한 새로운 컨텍스트 엔지니어링 규칙 (x.com/trq212)
- OpenAI와 Anthropic, 자사 수익을 위협하는 오픈 웨이트 AI에 공동 대응 (axios.com)
- 중국 AI 모델, 비용 경쟁력은 매력적…기업은 어디까지 활용해야 하나
- MS, 미스트랄과 손잡고 주권형 AI 승부수…기업 AI 선택권 넓힌다
- AI 기업은 왜 성공한 기술을 공짜로 내줄까
- AI 시대 SaaS 생존…소프트웨어는 기능을 팔고 플랫폼은 책임을 판다
- OECD, 육체 노동직도 AI 자동화 위험군으로 경고
- AI 모델 선택에 매몰되면 안 되는 이유
- “어떻게 만들고 어떻게 관리할 것인가” 데이터 제품 개발의 성패를 가르는 5가지 핵심 질문
- 🧾 홈택스 도움 스킬 (hometax-doum)
- Laguna XS 2.1 is a 33B total parameter Mixture-of-Experts model with 3B activated parameters per token designed for agentic coding and long-horizon work on a local machine.
- LM Studio Bionic: 오픈 모델을 위한 AI 에이전트 (lmstudio.ai)
- 합성 소비자, 실전 활용 가능성을 입증하다 (bain.com)
- ChatGPT에 광고하기 (ads.openai.com)
- Minimum Viable Baselines for Local LLM Inference: Thoughts on Model and Cache Quantization
- gemma-4-E2B-it-MLX-4bit - 4-bit quantized version of gemma-4-E2B-it using MLX, optimized for Apple Silicon.
- ChatGPT, 개인 건강 데이터를 일반 대화에 통합하는 Health 출시 (openai.com)
- Open Weights and American AI Leadership
- Core ML vs. Core AI: Who Runs the Loop?
- 챗봇은 당신이 생각하는 것 이상으로 많은 걸 알아낸다. 결국 그것은 어디에 사용될까
- Reddit, 연 900억원($60m) 계약 만료 앞두고 Google AI 접근 차단 검토 (finance.yahoo.com)
- Physical AI trades lab coat for a hard hat
- AI 신탁에 길들여져 가는 사람들
- GigaToken - 언어 모델 토큰화를 약 1,000배 가속 (github.com/marcelroed)
- AI 에이전트 상거래 시대 연다…리눅스 재단, x402 재단 출범
- Announcing the Monetization Gateway: charge for any resource behind Cloudflare via x402
- 뉴욕주, AI 데이터센터 허가 일시 중단…전력망·지역사회 보호 나선다
- OpenWorker is an open-source AI coworker that lives on your desktop and delivers finished work, not just chat: a polished document, a Slack reply with the numbers, an updated calendar, a triaged inbox.
- Best models for your hardware, GPU (July 2026): If you’re running local models on limited VRAM🔥
- AI의 거시경제 영향 분석 by KDI
- Fara1.5-27B is a multimodal computer use agent (CUA) for web browsers, from Microsoft Research AI Frontiers.
- Want to master LLM Cache Management? Start with these sources.
- 🤯Laguna S-2.1 UD-IQ3_S on a single 3090
- Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge
- Nanbeige4.2-3B is a compact agentic model built on Nanbeige4.2-3B-Base, designed to combine strong agentic behavior with broad reasoning and alignment capabilities.
- Ternary Bonsai 27B — GGUF: Full 27B-class reasoning in ternary transformer weights, for llama.cpp (CUDA, Metal, CPU)
- AI 연구소들은 ‘자전거 타는 펠리컨’에 맞춰 모델을 최적화했을까? (dylancastillo.co)
- Kimi K3는 Fable과 경쟁하며, 두 모델의 조합은 최고 수준의 성능을 달성함 (fireworks.ai)
- OmniRoute — 흩어진 무료/저가 AI 티어를 하나로 묶는 게이트웨이 (github.com/diegosouzapw)
- 그저 ‘적당한’ 세계가 온다 [이대한의 낭만연구실]
- Nanbeige4.2-3B is a compact agentic model built on Nanbeige4.2-3B-Base, designed to combine strong agentic behavior with broad reasoning and alignment capabilities.
- ultraprompt - A project to distill the reasoning traces of a frontier coding model into portable strategy skills for other agents.
- Qwen3.5-9B-Uncensored-HauhauCS-Aggressive
- OpenPlanter - A recursive-language-model investigation agent with a desktop GUI and terminal interface.
- VulnLLM-R-7B: Specialized Reasoning LLM for Vulnerability Detection
- Qwen-Image-3.0 - 풍부한 콘텐츠, 사실적 디테일, 깊이 있는 지식 (qwen.ai)
- AI는 왜 존재하지 않는 출처까지 만들어낼까
- AI 시대의 데이터 관리 (substack.com/williaminmon)
- 구글 제미나이 새 모델(3.6 Flash, 3.5 Flash-Lite, 3.5 Flash Cyber) 출시 (blog.google)
- JamAI Base is an open-source RAG (Retrieval-Augmented Generation) backend platform that integrates an embedded database (SQLite) and an embedded vector database (LanceDB) with managed memory and RAG capabilities.
- 지피티(클로드 등등)로 법률 자문 하는 방법 알려드림⚖️
- oMLX is the most convenient way to run MLX models on your Mac, and the fastest way to run them, with custom Metal kernels for GLM-5.2, MiniMax M3, DeepSeek V4, and Qwen3.5/3.6.
- graphify - Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph.
- Kimi K3·Qwen 3.8과 Anthropic의 잠재적 균열 (emergingtrajectories.com)
- The cleanup trap: Stop asking RAG to fix bad data
- Kimi Work: 지식 노동자를 위한 차세대 데스크톱 AI 에이전트 (kimi.com)
- 중국의 오픈 가중치 AI 전략이 앞서고 있음 (werd.io)
- 누가 중국 모델을 두려워하는가? (stratechery.com)
- SkillOpt: Executive Strategy for Self-Evolving Agent Skills
- Auto Company: A fully autonomous AI company running 24/7
- Train & run models on AMD GPUs with Unsloth
- AI 산업 이제 정유산업 비슷해져 간다: 기반 시설인 데이터 센터 구축에 사활
- Moonshot AI, Kimi K3 수요 급증으로 신규 구독 일시 중단 (twitter.com/kimi_moonshot)
- PIKE-RAG: 전문 지식과 추론으로 정확도를 높인 Microsoft의 산업용 RAG 프레임워크
- NuMarkdown-8B-Thinking is the first reasoning OCR VLM. It is specifically trained to convert documents into clean Markdown files, well suited for RAG applications.
- light-ocr - Fast, offline OCR for Node.js and C++.
- Gemma4-12B-Coder (GGUF) — Composer 2.5 × Fable 5 ✨
- Qwen 3.8 출시 (twitter.com/Alibaba_Qwen)
- From Brain Waves to Words: Brain2Qwerty Offers a New Path to Communication Without Surgery
- 기술적으로 뛰어난 데이터 팀이 여전히 실패하는 이유 (substack.com/practicaldatacommunity)
- 지금은 Kimi K3의 시간 (stephen.bochinski.dev)
- 데이터가 유일한 해자다 (substack.com/frontierai)
- 태스크 이코노미 - 데이터가 만드는 다음 1조 달러 시장 (x.com/EverettRandle)
- Chandra OCR 2 is a state of the art OCR model that converts images and PDFs into structured HTML/Markdown/JSON while preserving layout information.
- AINews] Microsoft Build: MAI-Thinking-1 and MAI Family models
- MiMo-V2.5-Pro-UltraSpeed: Pushing 1T-Parameter Model Generation Speed to 1000 TPS
- TileRT: Tile-Based Runtime for Ultra-Low-Latency LLM Inference
- Google announces Gemini 3.5 Live Translate for instant voice-to-voice translation
- Open R1 - A fully open reproduction of DeepSeek-R1. This repo is a work in progress, let's build it together!
- GLM-5.2 is the new leading open weights model on the Artificial Analysis Intelligence Index
- VibeThinker-3B: Exploring the Frontier of Verifiable Reasoning in Small Language Models
- VibeThinker-3B is a further exploration of the VibeThinker series at the 3B-parameter scale, focusing on challenging reasoning tasks with clear verification signals, such as mathematics, coding, and STEM.
- With Foundry, Microsoft bets the enterprise AI battle is about reliability, not capability
- LLM Visualization: 3D interactive model of a GPT-style LLM network running inference.
- 7 Top Mixture of Experts AI Models for Developers in 2026
- The Open-Source LLM Landscape in 2026
- Ontology Playground (Preview) ☕
- Kimi K3와 펠리컨 벤치마크에서 여전히 배울 수 있는 것 (simonwillison.net)
- 문어가 보여주는 지능의 진화 경로
- AGI 평가 과정과 수상자 선정에서 드러난 불일치 (kaggle.com)
- Claude Fable 5, 7월 20일부터 Max/Team Premium 요금제에 기본 포함 (x.com/claudeai)
- 2026년 7월 오픈소스 AI 현황 (stateofopensource.ai)
- Fable 5: Guardrails and burn rate are annoying users, who say it’s still better than Opus 4.8
- Microsoft’s open-source SkillOpt automatically upgrades AI agent skills without touching model weights
- ChatGPT에서 락다운 모드 및 일관된 “높은 위험” 레이블 도입
- Recommendation algorithms might be making your entertainment boring, new research suggests
- 중국 AI가 미국을 ‘빠르게 따라잡는 것’과 ‘실제로 추월하는 것’의 차이.
- 앤트로픽, 클로드 AI의 블랙홀 속을 들여다보다
- 챗GPT 워크 출시의 의미…’비용 효율 경쟁’ 시대로 재편되는 기업용 AI 판도
- AGI 시대 안전기준 누가 만드나…딥마인드 CEO 제안에 “필요” vs “빅테크 셀프 규제”
- NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
- Kimi K3 공개 - 개방형 프론티어 인텔리전스 (kimi.com)
- The Commercial Landscape of AI Sovereignty Offerings
- Hack Reveals Suno AI Music Generator Scraped YouTube, Deezer, and Genius
- OpenAI의 전 CTO(최고기술책임자)이자 ChatGPT, GPT-4 개발을 이끌었던 Mira Murati가 설립한 AI 스타트업 Thinking Machines Lab이 첫 번째 AI 모델 'Inkling'을 공개
- Inkling is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs.
- 「Machine Learning Study 혼자 해보기」 (github.com/teddylee777)
- Inkling: Thinking Machines Lab의 오픈 웨이트 모델 (thinkingmachines.ai)
- ChatGPT는 실제로 어떻게 출처를 고르는가 (네트워크 트래픽 분석) (suganthan.com)
- PrismML Releases Bonsai 27B: 1-bit and Ternary Builds of Qwen3.6-27B That Run on Laptops and Phones
- 우리는 너무 많은 생각을 AI에 떠넘기고 있는가? (artfish.ai)
- We’re rolling out some big improvements to Gemma 4, fueled by incredible community feedback and contributions! Here is a breakdown of what’s being fixed and updated in this release: 🧵👇
- Bonsai 27B - 휴대폰에서 실행되는 27B급 모델 (prismml.com)
- 프로덕션 AI 에이전트를 GPT-5.6으로 전환해 2.2배 빠르고 27% 저렴해진 과정 (ploy.ai)
- Deploying quantized models on Amazon SageMaker AI with Unsloth
- 𝗟𝗠𝗖𝗮𝗰𝗵𝗲 𝗶𝘀 𝗻𝗼𝘄 𝗯𝘂𝗶𝗹𝘁 𝗶𝗻𝘁𝗼 𝘃𝗟𝗟𝗠'𝘀 𝗼𝗳𝗳𝗶𝗰𝗶𝗮𝗹 𝗿𝗲𝗰𝗶𝗽𝗲𝘀 🎉
- 데이터 품질에 관하여 - 기본 원리 (substack.com/pivotal)
- Apple SpeechAnalyzer API, Whisper·이전 API와 비교 벤치마크 (get-inscribe.com)
- 온라인 플랫폼과 서비스가 주주들의 단기 이익을 극대화하기 위해 시간이 지남에 따라 품질이 저하되는 현상을 묘사하기 위해 '엔쉬티피케이션enshittification'이라는 신조어를 만들어냈던 저자 Cory Doctorow의 신간 The Reverse Centaur’s Guide to Life After AI 리뷰.
- '역-정보의 역설(Reverse Information Paradox)'
- 숏폼 동영상이 B2B 검색 결과와 AI 답변으로 영역을 확장하고 있다 (foundationinc.co)
- AI 토큰은 데이터센터를 어떻게 여행하는가 (datagravity.dev)
- LLM은 사랑하지만 과대광고는 싫다 (geohot.github.io)
- S&P 500 기업들이 실제로 AI를 어떻게 수용하고 이행하고 있는지를 순위로 매긴 'AIDE 인덱스'가 최근 발표되었음.
- "여러 모델 섞어 써도 한계 명확"…통념 깨진 '오케스트레이션' 효과
- 하드웨어
- 스팀 머신 리뷰 : 귀엽지만 아쉬운 161만원짜리 미니 PC
- 차는 살아남도록 설계됐다. 그런데 구조되도록 설계는 되었는가?
- 로드맵: AI 데이터센터 스택 (bvp.com)
- 나를 설레게 하는 차분한 기술들 (abhi.now)
- 모든 개발자가 SIMD를 알아야 하는 이유 (mitchellh.com)
- 삼성전자, CEO 직속 ‘RX사업추진실’ 출범…현대차 로봇 전략 이끈 이동건 부사장 선임
- MIDI 레코더 2,500대를 판매하며 배운 것: 하드웨어는 그렇게 어렵지 않다 (chipweinberger.com)
- Hotter Than a Hot Tub: The 45°C Breakthrough to Cool AI’s Biggest Machines
- Apple, OpenAI 직원 수십 명에게 법적 서한 발송 (ft.com)
- 소프트웨어가 세상을 먹어 치웠고, 이제 하드웨어가 소프트웨어를 먹고 있다 (wing.vc)
- 나는 USB-C 맥시멀리스트다 (shkspr.mobi)
- ROCm 7.14: TheRock Goes Production and Expands AMD’s AI Software Platform
- TrustMotion: SDV의 시스템 오케스트레이터
- DGX Spark vs AMD Strix Halo vs Mac Mini: The Honest 3-Way Benchmark
- Codex Micro (openai.com)
- 전 세계 스마트폰 출하량이 2026년 2분기 전년 대비 11% 감소하며 13년 만의 최저 수준으로 하락
- 목숨 걸고 타는 자율주행 차에서 컴퓨터가 갑자기 렉 걸리거나 블루스크린이 뜨면 어떻게 될까?
- Apple Silicon 임원이 말한 Mac mini AI 수요와 온디바이스 미래 (macrumors.com)
- 읽을거리 EOB

댓글 없음:
댓글 쓰기