Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#context-compression
Tag413건YouTube 25Article 388

#context-compression

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 295 · 연관도 61%#prompt-library공동문서 155 · 연관도 53%#retrieval-index공동문서 167 · 연관도 51%#applications공동문서 343 · 연관도 51%#semiconductors공동문서 360 · 연관도 51%#llm공동문서 349 · 연관도 47%#agent-routing공동문서 116 · 연관도 24%#ai-architecture공동문서 96 · 연관도 22%#agent-deployment공동문서 81 · 연관도 21%#compute공동문서 36 · 연관도 18%
Introducing Precursor: detecting agentic behavior with continuous client-side signals
Article2026년 7월 13일

Introducing Precursor: detecting agentic behavior with continuous client-side signals

Cloudflare의 Precursor는 웹 애플리케이션 전체 세션에서 최소한의 행동 신호를 지속적으로 수집·평가해 사람과 자동화·에이전트 트래픽을 구분하고, 정상 사용자의 불필요한 인증 마찰을 줄이는 클라이언트 측 검증 시스템이다.

blog.cloudflare.com
#privacy-design#semiconductors#applications#compute
OpenAI GPT-5.6 Sol, Terra, and Luna are now generally available on Amazon Bedrock
Article2026년 7월 13일

OpenAI GPT-5.6 Sol, Terra, and Luna are now generally available on Amazon Bedrock

OpenAI의 GPT 5.6 Sol·Terra·Luna가 Amazon Bedrock에서 정식 제공되며, 기업은 추론 성능·속도·비용에 맞춰 모델을 선택하고 확장형 추론, 프롬프트 캐싱, 데이터 보안 기능을 함께 활용할 수 있게 됐다.

aws.amazon.com
#openai#agent-routing#context-compression#prompt-library
Satya Nadella has issued a shocking warning to companies using AI
Article2026년 7월 13일

Satya Nadella has issued a shocking warning to companies using AI

사티아 나델라는 기업이 독점 AI 모델을 사용할 때 이용료뿐 아니라 프롬프트·피드백·업무 노하우까지 제공하게 된다며, 데이터 소유권을 유지하고 여러 모델을 전환할 수 있는 자체 학습 환경을 구축해야 한다고 경고했다.

Julie Bort
#anthropic#context-compression#prompt-library#llm
Recursive Knowledge Calibration Systems
Article2026년 7월 12일

Recursive Knowledge Calibration Systems

재귀적 지식 보정 시스템은 AI가 새로운 정보와 피드백을 기존 지식과 반복적으로 비교·조정하여 더 정확하고 맥락에 맞는 결과를 제공하도록 설명하는 개념적 틀이다.

MA Research Collectives
#privacy-design#ai-ai#llm#semiconductors
Turning News Headlines Into Trading Signals Using NLP Models
Article2026년 7월 12일

Turning News Headlines Into Trading Signals Using NLP Models

금융 뉴스의 문맥·감성·대상·주제·시점을 수치화하고 시장 데이터 및 기대치와 결합하면, 헤드라인을 검증 가능한 정량 거래 신호의 입력으로 전환할 수 있다.

wire.insiderfinance.io
#privacy-design#llm#semiconductors#applications
OpenAI bets on families as ChatGPT goes deeper into households
Article2026년 7월 11일

OpenAI bets on families as ChatGPT goes deeper into households

OpenAI는 ChatGPT 이용층이 부모와 중장년층으로 확대되자 가족·보호자·고령자를 위한 전담 제품 역할을 신설하고, 가정용 AI에 필요한 연령별 경험과 안전장치 강화에 나서고 있다.

Jagmeet Singh
#anthropic#openai#service-design#agent-memory
OpenWiki Brains: Proactive Memory for AI Agents
Article2026년 7월 10일

OpenWiki Brains: Proactive Memory for AI Agents

OpenWiki Brains는 이메일·문서·저장소·소셜 웹 등 여러 정보원에서 관심 맥락을 능동적으로 수집하고, 이를 자동 갱신되는 로컬 위키로 만들어 AI 에이전트의 지속적인 기억으로 제공하는 오픈소스 프레임워크다.

langchain.com
#agent-memory#agent-routing#context-compression#retrieval-index
A New Experiential Gallery Just Might Change Your Mind About AI Art
Article2026년 7월 10일

A New Experiential Gallery Just Might Change Your Mind About AI Art

리픽 아나돌의 데이터랜드는 동의받아 구축한 자연 데이터, 투명한 모델 설명, 생체 반응형 감각 경험을 결합해 인공지능 예술이 단순 생성물이 아니라 인간을 다시 발견하는 예술적 도구가 될 수 있음을 보여준다.

Miles Klee
#llm#semiconductors#applications#agent-deployment
Disaggregated prefill and decode for LLM inference on SageMaker HyperPod
Article2026년 7월 10일

Disaggregated prefill and decode for LLM inference on SageMaker HyperPod

긴 프롬프트의 사전 채우기와 토큰 생성을 별도 GPU 풀로 분리하고 KV 캐시를 EFA 기반 RDMA로 전달해, 대규모 LLM 스트리밍 추론의 첫 토큰 시간과 토큰 간 지연을 독립적으로 최적화하는 방법을 설명한다.

aws.amazon.com
#service-design#ai-architecture#agent-memory#context-compression
Fine-tune NVIDIA Nemotron 3 models with Amazon SageMaker AI serverless model customization
Article2026년 7월 10일

Fine-tune NVIDIA Nemotron 3 models with Amazon SageMaker AI serverless model customization

Amazon SageMaker AI의 서버리스 모델 맞춤화를 이용하면 인프라를 직접 관리하지 않고도 NVIDIA Nemotron 3 Nano와 Super를 SFT, RLVR, RLAIF 방식으로 기업 데이터와 업무에 특화할 수 있다.

aws.amazon.com
#nvidia#service-design#ai-architecture#capex-cycle
Profiling in PyTorch (Part 3): Attention is all you profile
Article2026년 7월 10일

Profiling in PyTorch (Part 3): Attention is all you profile

파이토치 프로파일러로 어텐션 구현을 비교한 결과, 제자리 마스킹은 불필요한 메모리 복사를 제거했지만 SDPA 수학 백엔드는 텐서 코어 미사용과 매 호출 마스크 생성으로 인해 단순 구현보다 3.7배 느렸다.

huggingface.co
#ai-architecture#agent-memory#capex-cycle#context-compression
Robot Dogs, Teslas, and Rescue Helicopters: The UN AI Summit Was a Lot
Article2026년 7월 10일

Robot Dogs, Teslas, and Rescue Helicopters: The UN AI Summit Was a Lot

유엔의 AI 포 굿 정상회의는 인류 문제 해결이라는 이상을 내세웠지만, 빅테크 의존과 컴퓨팅 격차, 인권의 기술적 집행, 느린 국제 합의라는 현실 속에서 AI의 발전 속도를 따라잡을 수 있는지 되물었다.

Chris Stokel-Walker
#resort-experience#ai-architecture#llm#semiconductors
SK Hynix raises $26.5B in the biggest foreign IPO in US history, is urged to build new US fabs
Article2026년 7월 10일

SK Hynix raises $26.5B in the biggest foreign IPO in US history, is urged to build new US fabs

SK하이닉스는 AI용 고대역폭 메모리 수요를 바탕으로 미국 증시에서 외국 기업 사상 최대인 265억 달러를 조달했으며, 미국 정부는 한국에 집중된 메모리 반도체 생산시설의 미국 이전·확대를 요구하고 있다.

Kate Park
#anthropic#nvidia#agent-memory#ai-infrastructure
A new way to reflect on how you use Claude
Article2026년 7월 9일

A new way to reflect on how you use Claude

Anthropic은 Claude 사용 기록을 시각화하고 목표와의 정렬 여부, AI 활용 역량, 휴식 습관을 점검할 수 있는 ‘리플렉션’ 기능을 베타로 공개했다.

anthropic.com
#anthropic#privacy-design#agent-memory#context-compression
Anthropic’s new Claude feature is quietly selling you on AI
Article2026년 7월 9일

Anthropic’s new Claude feature is quietly selling you on AI

앤트로픽의 ‘Claude Reflect’는 이용 습관을 보여주고 휴식을 권하는 분석 기능인 동시에, Claude가 일상 업무에 얼마나 깊이 자리 잡았는지 체감하게 해 이용자 유지와 서비스 의존도를 높이는 장치다.

Sarah Perez
#anthropic#agent-memory#agent-routing#context-compression
Can AI answer the $3 trillion question?
Article2026년 7월 9일

Can AI answer the $3 trillion question?

AI 인프라에 투입된 막대한 자본을 정당화하려면 업계가 3조 달러를 벌어야 하지만, 저가 모델 확산과 토큰 가격 하락이 투자금 회수와 거시경제의 위험 요인으로 떠오르고 있다.

Tim Fernholz
#anthropic#nvidia#token-efficiency#agent-memory
Elon Musk praises Mythos/Fable, promises not to 'cut off' Anthropic
Article2026년 7월 9일

Elon Musk praises Mythos/Fable, promises not to 'cut off' Anthropic

일론 머스크는 앤트로픽의 미토스·페이블 모델을 업계 최고라고 치켜세우며 경쟁사라도 스페이스엑스의 컴퓨팅 인프라에서 부당하게 배제하지 않겠다고 약속했지만, 양사의 대규모 계약에는 상업적 이익과 기술적 긴장도 함께 얽혀 있다.

Julie Bort
#anthropic#llm#semiconductors#applications
GPT-5.6 is now the preferred model in Microsoft 365 Copilot
Article2026년 7월 9일

GPT-5.6 is now the preferred model in Microsoft 365 Copilot

OpenAI의 최신 플래그십 모델 GPT 5.6이 Microsoft 365 Copilot의 기본 선호 모델로 도입되어 문서 작성, 데이터 분석, 프레젠테이션 제작과 부서 간 협업을 지원한다.

openai.com
#openai#privacy-design#context-compression#prompt-library
The 1X Neo Robot Has Freaky Fast Fingers
Article2026년 7월 9일

The 1X Neo Robot Has Freaky Fast Fingers

1X의 가정용 휴머노이드 로봇 네오는 인간 손에 가까운 자유도와 초고속 손가락을 갖췄지만, 원격 조작자를 집 안으로 연결하는 구조 때문에 성능 검증과 사생활 보호라는 과제를 동시에 안고 있다.

Boone Ashworth
#ai-safety#semiconductors#applications#compute
Tuning the harness, not the model: a Nemotron 3 Ultra playbook
Article2026년 7월 8일

Tuning the harness, not the model: a Nemotron 3 Ultra playbook

모델 가중치와 생성 설정을 그대로 유지한 채 프롬프트·도구 설명·미들웨어를 평가 기반으로 조율해, Nemotron 3 Ultra의 에이전트 성능을 낮은 비용으로 최전선 모델에 근접시킨 실전 사례다.

langchain.com
#service-design#ai-architecture#context-compression#prompt-library
Posts by Ian Macartney
Article2026년 7월 8일

Posts by Ian Macartney

이 페이지는 Convex 개발자 경험을 담당하는 Ian Macartney의 글 목록으로, 인증·권한, 내구성 있는 워크플로, 협업 편집, 운영 성숙도, 확장성, AI 애플리케이션, 검증·세션·마이그레이션 같은 Convex 기반 풀스택 개발 주제를 폭넓게 정리한다.

stack.convex.dev
#service-design#agent-memory#agent-routing#context-compression
Trends in Artificial Intelligence
Article2026년 7월 8일

Trends in Artificial Intelligence

Epoch AI의 대시보드는 AI 발전이 추론 비용 하락, 훈련 컴퓨트 확대, 소프트웨어 효율 개선, AI 칩·데이터센터 확장, 투자 증가가 함께 맞물리며 빠르게 진행되고 있음을 보여준다.

epoch.ai
#agent-memory#ai-safety#context-compression#retrieval-index
Epoch Capabilities Index
Article2026년 7월 8일

Epoch Capabilities Index

Epoch Capabilities Index는 여러 AI 벤치마크 점수를 하나의 일반 능력 척도로 결합해 모델 간 비교를 돕는 Epoch AI의 지표입니다.

epoch.ai
#llm#semiconductors#applications#agent-memory
Hot French startup ZML releases free product to speed inference across lots of AI chips
Article2026년 7월 8일

Hot French startup ZML releases free product to speed inference across lots of AI chips

프랑스 AI 스타트업 ZML이 여러 종류의 AI 칩에서 오픈소스 대형언어모델 추론을 더 빠르고 유연하게 실행하도록 돕는 무료 제품 ZML/LLMD를 출시했다.

Anna Heim
#nvidia#ai-architecture#ai-infrastructure#capex-cycle
이전123…181 / 18다음