최근 글 — 59편 Recent Logs — 59
-
공격도 AI, 방어도 AI — 그런데 방어 AI가 방어를 막았다When Both Sides Use AI: Hugging Face Discloses the First Autonomous Agent Security Breach
2026년 7월, AI 에이전트가 처음부터 끝까지 수행한 해킹 사건이 공식 공개됐어요. 공격도 AI였고, 방어도 AI였는데 — 방어 AI가 포렌식 작업을 막는 일이 실제로 일어났어요.In July 2026, Hugging Face disclosed a breach where an autonomous AI agent executed the entire attack end-to-end — and discovered that commercial AI models' safety guardrails were actively blocking their own forensic response.
-
NVIDIA Jetson Thor T3000·T2000 — Blackwell이 양산 로봇에 들어간다Blackwell for the Edge: NVIDIA's Jetson Thor T3000 and T2000 Bring Physical AI to the Mainstream
NVIDIA가 Blackwell GPU 탑재 Jetson Thor 모듈 2종을 공개했어요. T3000은 865 FP4 TFLOPS, T2000은 400 FP4 TFLOPS — 연구소용이 아니라 양산 로보틱스용 하드웨어예요.NVIDIA unveiled two new Jetson Thor modules with Blackwell GPUs — the T3000 at 865 FP4 TFLOPS and T2000 at 400 FP4 TFLOPS — designed for mainstream robotics production, not research labs.
-
비용 20분의 1, 선도 클로즈드 모델 성능 — 6개 기업이 보여준 오픈 모델의 현재20x Lower Cost, Frontier-Grade Performance — What Six Companies Built With Open Models
Harvey는 법률 AI를 선도 클로즈드 모델 수준으로 만들었고, 비용은 런당 10분의 1이에요. NVIDIA가 오픈 모델 커스터마이징 사례 6개를 공개하면서 'AI 경쟁력은 어떤 모델보다 어떻게 쌓느냐에서 온다'고 주장해요.Harvey matched leading closed models on complex legal tasks at 10x lower cost per run. NVIDIA published six open model customization cases making the case that competitive AI advantage comes from how you build, not which model you pick.
-
벤치마크가 숨긴 음성 AI의 빈틈 — 40개 모델 100만 건 평가로 드러난 것들What Voice AI Benchmarks Are Hiding — Four Findings From 1 Million Human Evaluations
Hume AI가 40개 이상의 음성 모델을 100만 건 인간 평가로 측정했어요. 소음 속 오류율 4배, 어떤 모델도 모든 역량에서 상위 5위 안에 없어요. 현재 벤치마크가 실제 대화 품질을 과대평가하는 이유를 살펴봤어요.Hume AI evaluated 40+ voice AI models with 1M+ human ratings. Noise drove 4x higher error rates; no single model topped all capability groups. Here's what the existing benchmarks are missing.
-
Claude for Teachers 출시 — Anthropic은 왜 학생이 아니라 교사를 골랐을까Claude for Teachers — Why Anthropic Bet on Teachers, Not Students
Anthropic이 미국 K-12 인증 교사에게 premium Claude를 1년 무료로 제공하는 Claude for Teachers를 출시했어요. 50개 주 학업 표준 정렬, 교육 도구 9종 통합, 교원노조와의 사전 협력까지 — 교육 시장 진입 전략을 뜯어봤습니다.Anthropic launched Claude for Teachers, giving verified US K-12 educators a free year of premium Claude. Standards alignment across all 50 states, nine edtech integrations, and a union partnership — a close look at the education play.
-
Anthropic이 캐나다 연구기관 8곳에 1,000만 캐나다달러를 넣었어요 — 다음 경쟁은 연구 생태계입니다Anthropic Puts CA$10M Into Canadian Research — The Next Frontier-Lab Race Is Ecosystems
Anthropic이 캐나다 AI 연구소·병원·대학 8곳에 1,000만 캐나다달러 규모의 Claude 크레딧 지원을 발표했어요. 모델 경쟁 다음 단계인 연구 생태계 선점 경쟁의 신호로 읽히는 이유를 정리했습니다.Anthropic announced CA$10M in Claude credits for eight Canadian research institutions. Here's why this looks like the opening move in a new race: not better models, but research ecosystems.
-
지시 한 번으로 완성까지 — ChatGPT Work가 바꾸는 업무 에이전트의 기준From Chat to Completion — How ChatGPT Work Redefines the AI Work Partner
ChatGPT Work는 초안을 주는 게 아니라 슬라이드·스프레드시트·웹앱을 수 시간에 걸쳐 직접 완성해주는 에이전트예요. 에이전트 AI가 실제 업무 흐름에 어떻게 진입하는지 살펴봤어요.ChatGPT Work is OpenAI's shift from chat assistant to task executor — finishing slides, docs, and web apps over hours with Scheduled Tasks and Computer Use built in.
-
코딩·에이전트에 최적화, $2 입력 가격 — xAI Grok 4.5가 고른 자리Trained Alongside Cursor, Priced at $2 — How xAI Is Positioning Grok 4.5
Grok 4.5는 Cursor와 함께 훈련한 코딩·에이전트 특화 모델로, 80 TPS 속도와 $2/1M 입력 가격을 내세워요. 프런티어 3파전이 벌어진 한 주에 xAI가 선택한 자리를 살펴봤어요.Grok 4.5 is xAI's coding and agentic tasks specialist, trained alongside Cursor, running at 80 TPS with a $2/1M input price — a deliberate position in a week with three frontier launches.
-
AI의 경제적 영향을 감시할 사람 — Ben Bernanke, Anthropic LTBT에 합류했어요Who Watches for AI's Economic Impact — Ben Bernanke Joins Anthropic's LTBT
노벨 경제학상 수상자이자 前 연방준비제도 의장 Ben Bernanke가 Anthropic의 장기편익신탁(LTBT)에 합류했어요. AI 거버넌스 조직이 어떤 역할을 하는지, 왜 경제학자가 필요한지 살펴봤어요.Nobel laureate and former Fed Chair Ben Bernanke has joined Anthropic's Long-Term Benefit Trust, the independent body that can appoint Anthropic's board. Here's what the LTBT does and why this appointment matters.
-
프롬프트를 코드처럼 — Mistral Studio에 버전 관리·감사 로그가 생겼어요Prompts as Production Assets — Mistral Studio Now Has Version Control and Audit Trails
Mistral Studio에 프롬프트·스킬의 버전 관리, 감사 로그, MCP 서버 직접 연동 기능이 추가됐어요. AI를 기업 운영에 내재화하려는 팀이 실제로 부딪히는 문제를 겨냥한 업데이트예요.Mistral Studio added immutable versioning, audit logs, environment labels, and direct MCP integration for prompts and skills. It's a product answer to the question enterprise AI teams keep asking: which version is actually running in production?