Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#context-compression
Tag404건YouTube 25Article 379

#context-compression

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 286 · 연관도 60%#prompt-library공동문서 154 · 연관도 53%#semiconductors공동문서 354 · 연관도 50%#applications공동문서 334 · 연관도 50%#retrieval-index공동문서 160 · 연관도 49%#llm공동문서 342 · 연관도 47%#agent-routing공동문서 116 · 연관도 25%#agent-deployment공동문서 80 · 연관도 21%#ai-architecture공동문서 92 · 연관도 21%#compute공동문서 34 · 연관도 17%
How to Give AI Agents a Bash Terminal Without Docker or VMs
Article2026년 5월 14일

How to Give AI Agents a Bash Terminal Without Docker or VMs

Convex Sandbox는 Docker나 VM 없이 Convex Node action, just bash, Convex storage를 조합해 AI 에이전트에게 상태가 유지되는 bash 터미널과 가상 파일 시스템을 제공하는 경량 샌드박스입니다.

stack.convex.dev
#agent-memory#agent-routing#capex-cycle#context-compression
Unlocking asynchronicity in continuous batching
Article2026년 5월 14일

Unlocking asynchronicity in continuous batching

연속 배치의 동기식 유휴 시간을 없애기 위해 중앙처리장치의 배치 준비와 그래픽처리장치의 계산을 분리하고, 비기본 CUDA 스트림과 이벤트로 병렬 실행과 작업 순서를 함께 보장하는 방법을 설명한다.

huggingface.co
#agent-routing#llm#semiconductors#applications
Introducing Langsmith Engine
Article2026년 5월 13일

Introducing Langsmith Engine

LangSmith Engine은 운영 트레이스에서 반복되는 에이전트 실패를 자동으로 묶고, 원인을 진단하며, 수정 PR과 평가 커버리지를 제안하는 LangSmith의 공개 베타 기능입니다.

langchain.com
#agent-routing#context-compression#prompt-library#semiconductors
Pinecone Just Demoted Vector Search. Here''''s the Knowledge Layer.
YouTube2026년 5월 13일

Pinecone Just Demoted Vector Search. Here''''s the Knowledge Layer.

벡터 검색은 에이전트 메모리의 한 부품일 뿐이며, 프로덕션 에이전트에는 작업별 데이터 계약·권한·출처·구조·관계를 함께 다루는 “지식 레이어”가 필요합니다.

AI News & Strategy Daily
#demoted#here#ai-distribution#context-compression
How ‘learnrights’ would compensate creators for AI model training
Article2026년 5월 12일

How ‘learnrights’ would compensate creators for AI model training

MIT 슬론의 토머스 말론과 공동 저자들은 생성형 AI 학습에 쓰이는 저작물에 대해 창작자가 라이선스 권리를 갖고 보상받도록 하는 ‘learnright’ 제도를 제안한다.

mitsloan.mit.edu
#anthropic#privacy-design#llm#semiconductors
Introducing Question and Highlights: High-Quality Answers from the Web, 100x Fewer Tokens
Article2026년 5월 8일

Introducing Question and Highlights: High-Quality Answers from the Web, 100x Fewer Tokens

Firecrawl은 /scrape에 question과 highlights 형식을 추가해, 전체 페이지를 긁어 LLM에 넣는 복잡한 과정을 한 번의 호출로 줄이고 URL 기반의 근거 있는 답변이나 원문 그대로의 발췌를 훨씬 적은 토큰으로 반환한다고 발표했다.

Eric Ciarla
#token-efficiency#agent-routing#context-compression#prompt-library
Running Codex safely at OpenAI
Article2026년 5월 8일

Running Codex safely at OpenAI

OpenAI는 Codex를 제한된 실행 환경, 승인 정책, 관리형 네트워크·인증 통제, 에이전트 친화적 텔레메트리로 운영해 개발 생산성과 보안 가시성을 함께 확보한다고 설명한다.

openai.com
#openai#privacy-design#agent-routing#context-compression
Testing ads in ChatGPT
Article2026년 5월 7일

Testing ads in ChatGPT

OpenAI는 ChatGPT의 무료·저가 접근성을 유지하기 위해 광고 파일럿을 확대하되, 답변 독립성·대화 프라이버시·사용자 통제를 핵심 원칙으로 유지한다고 밝혔다.

openai.com
#openai#privacy-design#agent-memory#context-compression
Introducing ChatGPT Futures: Class of 2026
Article2026년 5월 6일

Introducing ChatGPT Futures: Class of 2026

OpenAI는 AI를 책임감 있고 창의적으로 활용해 실제 변화를 만들고 있는 학생·젊은 빌더 26명을 ‘ChatGPT Futures: Class of 2026’ 첫 기수로 소개했다.

openai.com
#openai#privacy-design#context-compression#prompt-library
vLLM V0 to V1: Correctness Before Corrections in RL
Article2026년 5월 6일

vLLM V0 to V1: Correctness Before Corrections in RL

온라인 강화학습의 vLLM V0→V1 전환에서는 목적함수 보정을 서두르기보다 로그확률 의미, 런타임 기본값, 비행 중 가중치 갱신, fp32 출력 헤드를 먼저 일치시켜 추론 백엔드의 정확성을 복원해야 했다.

huggingface.co
#llm#semiconductors#applications#agent-deployment
GPT-5.5 Instant: smarter, clearer, and more personalized
Article2026년 5월 5일

GPT-5.5 Instant: smarter, clearer, and more personalized

OpenAI는 GPT 5.5 Instant를 ChatGPT의 기본 모델로 업데이트해 더 정확하고 간결하며 개인 맥락에 맞는 답변을 제공한다고 발표했다.

openai.com
#openai#privacy-design#context-compression#prompt-library
Stanford Merges AI and Data Science Efforts Under Single Institute
Article2026년 5월 4일

Stanford Merges AI and Data Science Efforts Under Single Institute

스탠퍼드는 인간중심 AI 연구소와 데이터 사이언스 이니셔티브를 Stanford HAI 이름 아래 통합하고, 제임스 랜데이가 이끄는 개방적·인간중심 AI·데이터 과학 거점으로 재편한다.

hai.stanford.edu
#llm#semiconductors#applications#agent-deployment
Lockdown Mode: /scrape Without Touching the Web
Article2026년 4월 30일

Lockdown Mode: /scrape Without Touching the Web

Firecrawl의 Lockdown Mode는 /scrape 요청을 기존 캐시 인덱스에서만 처리해 외부 웹 요청과 데이터 보존을 차단하는 보안 중심 스크래핑 모드다.

Eric Ciarla
#agent-routing#context-compression#prompt-library#llm
Transcript: ‘How Stripe Is Building for an Agent-native World’
Article2026년 4월 29일

Transcript: ‘How Stripe Is Building for an Agent-native World’

Stripe의 Emily Glassberg Sands는 인터넷 경제의 주체가 인간에서 AI 에이전트와 소프트웨어로 확장되면서 결제, 과금, 사기 탐지, 신원 인프라가 거래 순간이 아니라 고객 생애주기 전체를 다루도록 바뀌고 있다고 설명한다.

Dan Shipper
#anthropic#privacy-design#service-design#ai-architecture
Using AI at work: A practical 90-day guide
Article2026년 4월 28일

Using AI at work: A practical 90-day guide

이 글은 직장에서 AI를 막연히 두려워하거나 미루지 않고, 90일 동안 업무 분류·실험·인간 고유 역량 강화·커리어 재설계를 단계적으로 실행하는 실용적 가이드를 제시한다.

news.microsoft.com
#privacy-design#change-management#context-compression#prompt-library
Partnering with Ineffable Intelligence: A Superlearner for the Era of Experience
Article2026년 4월 27일

Partnering with Ineffable Intelligence: A Superlearner for the Era of Experience

세쿼이아는 데이비드 실버와 런던의 새 AI 연구소 Ineffable Intelligence에 투자하며, 인간 데이터 모방 없이 경험만으로 지식을 발견하는 강화학습 기반 ‘슈퍼러너’를 차세대 AI 경로로 제시한다.

sequoiacap.com
#llm#semiconductors#applications#change-management
It's all about the angle: Your photos, re-composed
Article2026년 4월 22일

It's all about the angle: Your photos, re-composed

Google Photos의 Auto frame은 촬영 후에도 사진을 3D 장면처럼 해석해 시점과 구도를 자동으로 재구성하는 이미지 편집 기능이다.

research.google
#llm#semiconductors#applications#agent-memory
QIMMA قِمّة ⛰: A Quality-First Arabic LLM Leaderboard
Article2026년 4월 21일

QIMMA قِمّة ⛰: A Quality-First Arabic LLM Leaderboard

QIMMA는 아랍어 LLM 평가에서 벤치마크를 먼저 검증한 뒤 모델을 평가함으로써, 점수가 실제 아랍어 능력을 더 신뢰성 있게 반영하도록 설계된 품질 우선 리더보드입니다.

huggingface.co
#agent-routing#workflow-automation#llm#semiconductors
Scaling Codex to enterprises worldwide
Article2026년 4월 21일

Scaling Codex to enterprises worldwide

Codex는 주간 개발자 사용자가 2주 만에 300만 명 이상에서 400만 명 이상으로 늘어난 가운데, 기업들이 소프트웨어 개발과 지식 업무 전반의 실제 워크플로에 빠르게 도입하는 단계로 확장되고 있다.

openai.com
#agent-memory#agent-routing#context-compression#retrieval-index
Designing synthetic datasets for the real world: Mechanism design and reasoning from first principles
Article2026년 4월 16일

Designing synthetic datasets for the real world: Mechanism design and reasoning from first principles

Simula는 합성 데이터 생성을 개별 샘플 제작이 아니라 데이터셋 전체의 범위, 난이도, 품질을 설계하는 메커니즘 디자인 문제로 재정의한 프레임워크다.

research.google
#privacy-design#agent-routing#context-compression#prompt-library
Meet HoloTab by HCompany. Your AI browser companion.
Article2026년 4월 15일

Meet HoloTab by HCompany. Your AI browser companion.

HoloTab은 사용자가 원하는 작업을 말하거나 한 번 시연하면 웹사이트 탐색과 반복 업무를 브라우저 안에서 대신 수행하는 무료 크롬 확장 프로그램이다.

huggingface.co
#llm#semiconductors#applications#agent-memory
Fragments: April 14
Article2026년 4월 14일

Fragments: April 14

마틴 파울러는 AI 시대에도 좋은 소프트웨어를 만드는 핵심은 더 많은 코드를 빠르게 생산하는 능력이 아니라, 인간의 ‘게으름’이 낳는 단순한 추상화와 의심·검증·절제라고 정리한다.

martinfowler.com
#agent-memory#retrieval-index#llm#semiconductors
Towards developing future-ready skills with generative AI
Article2026년 4월 13일

Towards developing future-ready skills with generative AI

구글 리서치는 생성형 AI 대화 시뮬레이션을 활용해 협업·갈등 해결·프로젝트 관리·창의성 같은 미래 준비 역량을 확장 가능하게 평가하는 연구 실험 Vantage를 공개했다.

research.google
#llm#semiconductors#applications#agent-memory
ConvApparel: Measuring and bridging the realism gap in user simulators
Article2026년 4월 9일

ConvApparel: Measuring and bridging the realism gap in user simulators

ConvApparel은 의류 쇼핑 대화 데이터를 통해 LLM 기반 사용자 시뮬레이터가 실제 인간과 얼마나 다른지 측정하고, 특히 예상 밖으로 나쁜 대화 에이전트에 적응하는지를 검증하는 데이터셋과 평가 프레임워크입니다.

research.google
#ai-distribution#context-compression#prompt-library#search-advertising
이전1…7891011…179 / 17다음