Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#context-compression
Tag404건YouTube 25Article 379

#context-compression

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 286 · 연관도 60%#prompt-library공동문서 154 · 연관도 53%#semiconductors공동문서 354 · 연관도 50%#applications공동문서 334 · 연관도 50%#retrieval-index공동문서 160 · 연관도 49%#llm공동문서 342 · 연관도 47%#agent-routing공동문서 116 · 연관도 25%#agent-deployment공동문서 80 · 연관도 21%#ai-architecture공동문서 92 · 연관도 21%#compute공동문서 34 · 연관도 17%
Introducing Spark 1 Pro and Spark 1 Mini
Article2026년 1월 14일

Introducing Spark 1 Pro and Spark 1 Mini

Firecrawl은 /agent 엔드포인트를 위해 비용 효율형 Spark 1 Mini와 고정확도형 Spark 1 Pro를 출시하며, 웹 데이터 추출 작업의 비용·정확도 선택지를 넓혔다.

Eric Ciarla
#anthropic#context-compression#prompt-library#llm
Dynamic surface codes open new avenues for quantum error correction
Article2026년 1월 13일

Dynamic surface codes open new avenues for quantum error correction

구글 퀀텀 AI는 Willow 프로세서에서 동적 표면 코드를 실험해 결합기 수를 줄이고, 누설로 인한 상관 오류를 억제하며, iSWAP 게이트 기반 오류 정정 가능성을 확인했다.

research.google
#privacy-design#ai-architecture#llm#semiconductors
Hard-braking events as indicators of road segment crash risk
Article2026년 1월 13일

Hard-braking events as indicators of road segment crash risk

구글 리서치는 안드로이드 오토에서 집계한 급제동 사건이 실제 도로 구간 사고율과 유의미한 양의 관계를 보이며, 희소한 사고 기록을 보완하는 도로 안전 선행 지표가 될 수 있음을 검증했다.

research.google
#privacy-design#llm#semiconductors#applications
Zenken boosts a lean sales team with ChatGPT Enterprise
Article2026년 1월 13일

Zenken boosts a lean sales team with ChatGPT Enterprise

일본 기업 젠켄은 챗GPT 엔터프라이즈를 전사적으로 도입해 영업 준비, 전략 분석, 번역, 문서 작성 업무를 줄이고 작은 팀으로도 매출 성장과 생산성 향상을 뒷받침하고 있다.

openai.com
#agent-routing#llm#semiconductors#applications
NeuralGCM harnesses AI to better simulate long-range global precipitation
Article2026년 1월 12일

NeuralGCM harnesses AI to better simulate long-range global precipitation

NeuralGCM은 물리 기반 대기 모델과 NASA 위성 강수 관측으로 학습한 신경망을 결합해 전 지구 강수, 특히 극한 강수와 하루 주기 강수의 장기 시뮬레이션 정확도를 높인 하이브리드 모델이다.

research.google
#llm#semiconductors#applications#agent-deployment
Cohere on Hugging Face Inference Providers 🔥
Article2026년 1월 9일

Cohere on Hugging Face Inference Providers 🔥

코히어가 허깅페이스 허브의 추론 제공자로 합류하면서 기업용 언어·검색·다국어·멀티모달 모델 9종을 웹 화면과 여러 클라이언트 라이브러리에서 서버리스 방식으로 사용할 수 있게 됐다.

huggingface.co
#multimodal#capex-cycle#llm#semiconductors
How Tolan builds voice-first AI with GPT-5.1
Article2026년 1월 7일

How Tolan builds voice-first AI with GPT-5.1

Tolan은 GPT 5.1의 낮은 지연시간과 높은 지시 이행 능력을 바탕으로, 매 턴 재구성되는 문맥과 정교한 기억·캐릭터 시스템을 결합해 자연스럽고 일관된 음성형 AI 동반자를 구현했다.

openai.com
#ai-architecture#multimodal#context-compression#prompt-library
Aligning to What? Rethinking Agent Generalization in MiniMax M2
Article2025년 12월 23일

Aligning to What? Rethinking Agent Generalization in MiniMax M2

MiniMax M2는 벤치마크 성능과 현실적 유용성을 함께 추구하면서, 도구 종류의 확대보다 전체 에이전트 실행 과정에서 발생하는 변화와 교란에 적응하는 능력을 일반화의 핵심으로 삼았다.

huggingface.co
#agent-routing#context-compression#m2-m2#prompt-library
Introducing /agent: Gather Data Wherever It Lives on the Web
Article2025년 12월 18일

Introducing /agent: Gather Data Wherever It Lives on the Web

Firecrawl의 /agent는 URL 지정이나 사이트별 스크래핑 코드 없이 프롬프트만으로 웹 전반을 검색·탐색·클릭하고 구조화된 데이터를 수집하도록 설계된 연구 프리뷰 기능입니다.

Eric Ciarla
#agent-routing#context-compression#prompt-library#llm
Measuring AI’s capability to accelerate biological research in the wet lab
Article2025년 12월 16일

Measuring AI’s capability to accelerate biological research in the wet lab

OpenAI와 Red Queen Bio는 GPT 5가 실제 습식 실험 반복 과정에서 분자 클로닝 프로토콜을 제안·분석·개선하도록 평가했고, 특정 모델 시스템에서 검증된 클론 회수 효율을 기준 대비 79배 높였다고 보고했다.

openai.com
#agent-routing#context-compression#prompt-library#llm
New in llama.cpp: Model Management
Article2025년 12월 11일

New in llama.cpp: Model Management

llama.cpp 서버의 라우터 모드는 재시작 없이 여러 GGUF 모델을 검색·적재·전환·해제하고, 요청의 모델 지정에 따라 독립 프로세스로 실행되는 모델에 트래픽을 전달합니다.

huggingface.co
#anthropic#ai-architecture#agent-memory#agent-routing
DeepMath: A lightweight math reasoning Agent with smolagents
Article2025년 12월 8일

DeepMath: A lightweight math reasoning Agent with smolagents

DeepMath는 Qwen3 4B Thinking 기반의 경량 수학 추론 에이전트로, 긴 사고 과정을 짧은 파이썬 실행 단계로 대체하고 GRPO 학습을 통해 더 간결하고 정확한 수학 풀이를 목표로 한다.

huggingface.co
#llm#semiconductors#applications#long-context
Titans + MIRAS: Helping AI have long-term memory
Article2025년 12월 4일

Titans + MIRAS: Helping AI have long-term memory

Titans와 MIRAS는 실행 중 장기 기억을 선택적으로 갱신하는 방식으로, 매우 긴 문맥을 더 빠르고 정확하게 다루려는 새로운 시퀀스 모델링 접근이다.

research.google
#ai-architecture#agent-memory#context-compression#retrieval-index
Anthropic’s Newest Model Blew This Founder’s Mind—And Made Him Uncomfortable
Article2025년 12월 3일

Anthropic’s Newest Model Blew This Founder’s Mind—And Made Him Uncomfortable

이 글은 폴 포드와 댄 시퍼의 대화를 통해 클로드 오퍼스 4.5가 코딩 경험을 크게 바꾸는 동시에, 프롬프트 편향·훈련 데이터 공개·직업 정체성의 흔들림 같은 불편한 질문을 드러냈다고 정리한다.

Rhea Purohit
#anthropic#privacy-design#change-management#context-compression
From Waveforms to Wisdom: The New Benchmark for Auditory Intelligence
Article2025년 12월 3일

From Waveforms to Wisdom: The New Benchmark for Auditory Intelligence

Google Research는 기계 청각 지능을 여덟 가지 핵심 능력으로 표준 평가하는 오픈소스 벤치마크 MSEB를 공개하며, 현재 사운드 임베딩 모델들이 범용성과 견고성에서 큰 성능 여지를 남기고 있다고 밝혔다.

research.google
#multimodal#llm#semiconductors#vision-language-models
Microsoft and Hugging Face expand collaboration
Article2025년 11월 24일

Microsoft and Hugging Face expand collaboration

마이크로소프트와 허깅페이스는 1만 개 이상의 검증된 공개 모델을 애저 AI 파운드리에서 간편하고 안전하게 배포할 수 있도록 협력을 확대했습니다.

huggingface.co
#agent-deployment#llm#semiconductors#applications
AI, networks and Mechanical Turks — Benedict Evans
Article2025년 11월 23일

AI, networks and Mechanical Turks — Benedict Evans

대형 소비자 인터넷 서비스는 사용자 행동의 상관관계로 추천과 발견을 만들어 왔지만, 대규모 언어 모델은 사물과 의도를 더 넓게 해석해 인터넷의 ‘발견 필터’ 자체를 바꿀 수 있다.

Benedict Evans
#llm#semiconductors#applications#agent-memory
One Year Since the “DeepSeek Moment”
Article2025년 11월 21일

One Year Since the “DeepSeek Moment”

딥시크 R1은 기술·도입·심리적 장벽을 낮춰 중국의 개방형 인공지능 생태계를 확장하고, 세계 경쟁의 중심을 개별 모델 성능에서 시스템·응용·생태계 역량으로 이동시킨 전환점이었다.

huggingface.co
#privacy-design#llm#semiconductors#applications
Reducing EV range anxiety: How a simple AI model predicts port availability
Article2025년 11월 21일

Reducing EV range anxiety: How a simple AI model predicts port availability

구글 리서치는 전기차 충전 대기와 주행거리 불안을 줄이기 위해, 특정 충전소에서 30~60분 뒤 포트가 비어 있을 가능성을 예측하는 경량 선형 회귀 모델을 개발·배포했다.

research.google
#privacy-design#agent-deployment#travel-hospitality#llm
How evals drive the next chapter in AI for businesses
Article2025년 11월 19일

How evals drive the next chapter in AI for businesses

기업용 인공지능의 성과를 높이려면 막연한 기대에 의존하지 말고, 조직의 실제 업무 맥락에서 성공 기준을 명시하고 측정하며 지속적으로 개선하는 평가 체계를 구축해야 한다.

openai.com
#agent-routing#context-compression#prompt-library#llm
Generative UI: A rich, custom, visual interactive user experience for any prompt
Article2025년 11월 18일

Generative UI: A rich, custom, visual interactive user experience for any prompt

구글은 사용자의 어떤 프롬프트에도 맞춰 웹페이지, 도구, 시뮬레이션 같은 맞춤형 인터랙티브 경험을 즉석 생성하는 생성형 UI 구현을 소개하고, Gemini 앱과 Google Search AI Mode에 실험적으로 적용한다고 밝혔다.

research.google
#multimodal#context-compression#prompt-library#search-advertising
A new quantum toolkit for optimization
Article2025년 11월 13일

A new quantum toolkit for optimization

Google Quantum AI 연구진은 최적화 문제를 디코딩 문제로 변환해 양자 간섭과 기존 디코딩 알고리즘을 결합하는 DQI 알고리즘을 제시하며, 특히 OPI 문제에서 알려진 고전 알고리즘 대비 뚜렷한 속도 우위를 보일 수 있음을 설명했다.

research.google
#privacy-design#semiconductors#applications#compute
Differentially private machine learning at scale with JAX-Privacy
Article2025년 11월 12일

Differentially private machine learning at scale with JAX-Privacy

Google은 JAX 기반의 차등 프라이버시 머신러닝 라이브러리 JAX Privacy 1.0을 공개하며, 대규모 모델 학습과 감사에 필요한 DP 알고리즘·스케일링·회계 도구를 통합했다고 발표했다.

research.google
#privacy-design#agent-routing#workflow-automation#llm
DS-STAR: A state-of-the-art versatile data science agent
Article2025년 11월 6일

DS-STAR: A state-of-the-art versatile data science agent

DS STAR는 다양한 데이터 형식 분석, LLM 기반 검증, 순차적 계획 개선을 결합해 복잡한 데이터 과학 작업을 자동화하고 주요 벤치마크에서 기존 에이전트보다 높은 성능을 보인 데이터 과학 에이전트입니다.

research.google
#agent-routing#llm#semiconductors#applications
이전1…1011121314…1712 / 17다음