Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#context-compression
Tag404건YouTube 25Article 379

#context-compression

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 286 · 연관도 60%#prompt-library공동문서 154 · 연관도 53%#semiconductors공동문서 354 · 연관도 50%#applications공동문서 334 · 연관도 50%#retrieval-index공동문서 160 · 연관도 49%#llm공동문서 342 · 연관도 47%#agent-routing공동문서 116 · 연관도 25%#agent-deployment공동문서 80 · 연관도 21%#ai-architecture공동문서 92 · 연관도 21%#compute공동문서 34 · 연관도 17%
Better decisions at scale: How mathematical optimization delivers where intuition fails
Article2026년 6월 8일

Better decisions at scale: How mathematical optimization delivers where intuition fails

수학적 최적화는 직관과 단순 규칙으로 처리하기 어려운 대규모 운영 의사결정을 제약 조건 안에서 검증 가능한 최적 해로 바꾸는 AI 접근법이다.

aws.amazon.com
#agent-routing#llm#semiconductors#applications
DeepSeek-V4: a million-token context that agents can actually use
Article2026년 6월 8일

DeepSeek-V4: a million-token context that agents can actually use

DeepSeek V4는 최고 벤치마크 점수보다 100만 토큰 문맥을 실제 에이전트 작업에서 감당하게 만드는 긴 문맥 효율, 도구 호출 지속성, 샌드박스 기반 학습 인프라에 초점을 둔 모델이다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
How to Build a Multi-Agent Workflow for LLM Wikis in Hermes Kanban
YouTube2026년 6월 8일

How to Build a Multi-Agent Workflow for LLM Wikis in Hermes Kanban

Hermes Kanban 기반 Multi Agent Workflow는 LLM 위키를 더 최신이고 검증 가능한 지식베이스로 유지하기 위한 실전형 자동화 구조다.

Tonbi''s AI Garage
#anthropic-model-roadmap#frontier-model-evaluation#agent-systems#core-thesis
Designing the hf CLI as an agent-optimized way to work with the Hub
Article2026년 6월 4일

Designing the hf CLI as an agent-optimized way to work with the Hub

Hugging Face는 hf CLI를 사람과 코딩 에이전트가 모두 효율적으로 Hub를 다룰 수 있도록 재설계했고, 특히 복잡한 다단계 작업에서 에이전트의 토큰 사용과 실패를 줄인다고 설명한다.

huggingface.co
#agent-routing#llm#semiconductors#applications
Stanford Robotics Seminar ENGR319
YouTube2026년 6월 4일

Stanford Robotics Seminar ENGR319

Geometry in Robot Learning의 핵심은 로봇 학습을 무작정 더 큰 데이터와 모델로 밀어붙이기보다, 기하·대칭성·좌표계 구조를 모델에 넣어 데이터 효율성과 자세 일반화를 높이려는 것이다.

Stanford Online
#anthropic-model-roadmap#frontier-model-evaluation#agent-systems#core-thesis
Direct Preference Optimization Beyond Chatbots
Article2026년 6월 3일

Direct Preference Optimization Beyond Chatbots

DharmaOCR는 SFT 모델이 스스로 생성한 반복 퇴화 출력을 DPO의 거부 사례로 활용해, 구조화 OCR에서 모든 실험 모델군의 텍스트 퇴화율을 낮췄다.

huggingface.co
#ai-architecture#multimodal#llm#semiconductors
How to Use the New Convex Codex Plugin in OpenAI Codex
Article2026년 6월 3일

How to Use the New Convex Codex Plugin in OpenAI Codex

Convex Codex 플러그인은 Codex 안에서 Convex 백엔드를 스캐폴딩·수정하게 해 주지만, 실제로 쓰려면 프롬프트에 반드시 @convex를 명시해야 한다.

stack.convex.dev
#openai#agent-deployment#agent-routing#context-compression
Reading Today’s Headlines Through AI: A Real-Time Audit of Six Commercial Chatbots
Article2026년 6월 3일

Reading Today’s Headlines Through AI: A Real-Time Audit of Six Commercial Chatbots

스탠퍼드 HAI 연구는 상용 AI 챗봇이 당일 뉴스 질문에서 높은 평균 정확도를 보였지만, 실제 신뢰성은 언어·지역별 검색 인프라, 출처 선택, 불완전한 질문에 대한 취약성에 크게 좌우된다고 분석했다.

hai.stanford.edu
#privacy-design#llm#semiconductors#applications
‘AI gravity’ is pulling you toward dependency. Here’s how to push back
Article2026년 6월 2일

‘AI gravity’ is pulling you toward dependency. Here’s how to push back

MIT Sloan의 에릭 소는 AI가 업무 효율을 높이는 동시에 인간의 사고·학습·조직 지식을 약화시키는 ‘AI gravity’를 만들 수 있다고 경고하며, 기업과 개인이 의도적으로 인지 역량을 보존해야 한다고 말한다.

mitsloan.mit.edu
#agent-routing#llm#semiconductors#applications
Heeding the pope’s call to ensure AI protects human dignity
Article2026년 6월 1일

Heeding the pope’s call to ensure AI protects human dignity

MIT Sloan의 Thomas A. Kochan은 AI 시대에 인간 존엄을 지키려면 노동자가 AI 도입과 활용 방향을 결정하는 과정에 실질적으로 참여하고, 기술이 만든 경제적 이익을 함께 나누는 새로운 사회계약이 필요하다고 주장한다.

mitsloan.mit.edu
#semiconductors#applications#compute#agent-memory
Compound Engineering Gets an Upgrade
Article2026년 5월 29일

Compound Engineering Gets an Upgrade

AI 모델이 더 유능해지면서 컴파운드 엔지니어링은 단순한 ‘계획 작업 검토’ 루프를 넘어, 인간이 시작과 끝에서 방향성과 품질을 책임지는 방식으로 확장되고 있다.

Kieran Klaassen
#every#kieran-klaassen#trevin-chow#compounding-feedback
April 2026: LangChain Newsletter
Article2026년 5월 28일

April 2026: LangChain Newsletter

2026년 4월 LangChain 뉴스레터는 LangSmith 평가·비용·도구 관리 업데이트, Deep Agents 배포 기능, Interrupt 2026 행사, 에이전트 개선 루프 콘텐츠와 고객 사례를 소개한다.

langchain.com
#multimodal#agent-deployment#agent-memory#context-compression
Reachy Mini goes fully local
Article2026년 5월 27일

Reachy Mini goes fully local

Reachy Mini의 음성 대화 전 과정을 로컬에서 실행하도록 VAD·STT·LLM·TTS 계단식 파이프라인을 구축하고, 필요에 따라 각 구성 요소와 추론 방식을 교체하는 방법을 설명한 안내서다.

huggingface.co
#privacy-design#llm#semiconductors#applications
AI Hiring Tools Can Yield Racial Bias and Systemic Rejection
Article2026년 5월 26일

AI Hiring Tools Can Yield Racial Bias and Systemic Rejection

스탠퍼드 HAI가 실제 채용 데이터를 대규모로 분석한 결과, 인공지능 채용 선별 도구가 직무별 인종적 불균형을 만들고 같은 지원자를 여러 채용 과정에서 반복적으로 배제할 수 있다는 우려가 제기됐다.

hai.stanford.edu
#privacy-design#llm#semiconductors#ai-coding
Firecrawl is now live on the Vercel Marketplace
Article2026년 5월 26일

Firecrawl is now live on the Vercel Marketplace

Firecrawl이 Vercel Marketplace의 네이티브 통합으로 출시되어, Vercel 프로젝트에서 웹 데이터 수집용 계정·API 키·청구 설정을 자동화할 수 있게 됐습니다.

Eric Ciarla
#agent-memory#agent-routing#context-compression#retrieval-index
On the Shifting Global Compute Landscape
Article2026년 5월 26일

On the Shifting Global Compute Landscape

미국 중심이던 인공지능 연산 생태계가 수출 통제, 중국산 칩의 성장, 개방형 모델과 연산 효율 기술의 확산을 계기로 중국을 포함한 다극적 하드웨어·소프트웨어 구조로 재편되고 있다.

huggingface.co
#nvidia#linear-attention#ai-architecture#agent-memory
Building Convex OS, a Browser-Based React App with Real-Time Sync
Article2026년 5월 23일

Building Convex OS, a Browser-Based React App with Real-Time Sync

Convex OS는 Windows XP 스타일의 브라우저 기반 React 데스크톱 UI를 Convex의 반응형 데이터베이스 상태 모델 위에 올려, 창 위치·프로세스·파일 상태를 여러 탭에서 실시간으로 공유하게 만든 실험이다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
All Systems Nominal: shaping the future of hardware development
Article2026년 5월 21일

All Systems Nominal: shaping the future of hardware development

전 해군 장교 캐머런 맥코드가 잠수함·의회·방산 스타트업 현장에서 겪은 복잡한 시스템 운용 경험은, 실제 배치되는 하드웨어를 더 빠르고 안전하게 시험하려는 Nominal의 문제의식으로 이어졌다.

sequoiacap.com
#service-design#llm#applications#robotics
New Approach to Scaling Laws Could Change How AI Models Are Trained
Article2026년 5월 21일

New Approach to Scaling Laws Could Change How AI Models Are Trained

스탠퍼드 연구진은 교육측정학의 문항응답 원리를 스케일링 법칙에 적용해, 대형 언어모델 성능 예측에 필요한 계산량을 크게 줄이는 아이템 응답 스케일링 법칙을 제안했다.

hai.stanford.edu
#privacy-design#ai-architecture#llm#semiconductors
Everything new in our Google AI subscriptions, fresh from I/O 2026
Article2026년 5월 19일

Everything new in our Google AI subscriptions, fresh from I/O 2026

Google은 I/O 2026에서 새로운 월 100달러 AI Ultra 요금제, 기존 최상위 Ultra 가격 인하, Gemini Omni·Gemini 3.5 Flash·Gemini Spark·Project Genie 접근 확대, Gmail·Gemini 앱 생산성 기능과 compute 기반 사용량 체계를 포함한 Google AI 구독 개편을 발표했다.

blog.google
#agent-routing#context-compression#prompt-library#search-advertising
OlmoEarth v1.1: A more efficient family of Earth observation models
Article2026년 5월 19일

OlmoEarth v1.1: A more efficient family of Earth observation models

OlmoEarth v1.1은 위성영상 토큰의 해상도별 분리를 통합하고 사전학습 방식을 조정해 기존 성능을 대체로 유지하면서 연산 비용을 최대 3분의 1로 줄인 지구관측 모델 제품군이다.

huggingface.co
#ai-architecture#llm#semiconductors#applications
Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL
Article2026년 5월 19일

Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL

TRL의 델타 가중치 동기화는 연속된 강화학습 단계에서 실제로 바뀐 극소수의 bf16 원소만 희소 세이프텐서로 저장해 허깅페이스 버킷으로 전달함으로써, 대규모 모델의 반복적인 전체 체크포인트 전송 비용과 추론 중단 시간을 크게 줄이는 방식이다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
At aged care provider Regis, AI takes on paperwork so staff can focus on residents
Article2026년 5월 18일

At aged care provider Regis, AI takes on paperwork so staff can focus on residents

호주 노인요양기관 Regis는 RegiCare Assist로 방대한 24시간 보고서와 인수인계 기록을 요약해 임상관리자가 서류보다 입소자 돌봄에 더 많은 시간을 쓰도록 하고 있습니다.

news.microsoft.com
#privacy-design#agent-memory#context-compression#retrieval-index
Resilient AI End-to-End Tests with Stagehand and Convex
Article2026년 5월 16일

Resilient AI End-to-End Tests with Stagehand and Convex

이 글은 Stagehand의 자연어 브라우저 제어와 Convex의 실행별 임시 백엔드를 결합해, 선택자 중심 E2E 테스트의 취약성을 줄이고 실제 사용자 의도에 가까운 테스트를 구성한 경험을 정리한다.

stack.convex.dev
#agent-deployment#agent-routing#llm#semiconductors
이전1…678910…178 / 17다음