Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#context-compression
Tag404건YouTube 25Article 379

#context-compression

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 286 · 연관도 60%#prompt-library공동문서 154 · 연관도 53%#semiconductors공동문서 354 · 연관도 50%#applications공동문서 334 · 연관도 50%#retrieval-index공동문서 160 · 연관도 49%#llm공동문서 342 · 연관도 47%#agent-routing공동문서 116 · 연관도 25%#agent-deployment공동문서 80 · 연관도 21%#ai-architecture공동문서 92 · 연관도 21%#compute공동문서 34 · 연관도 17%
Generate Images with Claude and Hugging Face
Article2025년 8월 18일

Generate Images with Claude and Hugging Face

Claude에 Hugging Face MCP 서버와 이미지 생성용 Gradio Space를 연결하면 프롬프트 작성, 결과 확인, 반복 개선을 대화 안에서 수행하며 Krea와 Qwen Image 같은 최신 모델을 활용할 수 있다.

huggingface.co
#anthropic#agent-routing#context-compression#prompt-library
Vision Language Model Alignment in TRL ⚡️
Article2025년 8월 16일

Vision Language Model Alignment in TRL ⚡️

TRL은 비전 언어 모델 정렬을 기존 SFT·DPO에서 MPO·GRPO·GSPO와 RLOO·Online DPO까지 확장해, 멀티모달 선호 학습과 강화학습을 더 유연하게 적용할 수 있도록 지원한다.

huggingface.co
#multimodal#capex-cycle#llm#semiconductors
Beyond billion-parameter burdens: Unlocking data synthesis with a conditional generator
Article2025년 8월 14일

Beyond billion-parameter burdens: Unlocking data synthesis with a conditional generator

Google Research는 거대 LLM을 비공개 데이터에 직접 미세조정하지 않고도, 1억4000만 매개변수 조건부 생성기와 주제 분포 매칭을 통해 차등 프라이버시 합성 텍스트 데이터를 만드는 CTCL 프레임워크를 제안했다.

research.google
#privacy-design#context-compression#prompt-library#llm
TextQuests: How Good are LLMs at Text-Based Video Games?
Article2025년 8월 13일

TextQuests: How Good are LLMs at Text-Based Video Games?

TextQuests는 25개의 고전 텍스트 어드벤처 게임을 통해 LLM 에이전트의 장기 문맥 추론, 탐색 학습, 공간 이해, 행동 효율성과 유해 행동 경향을 평가하는 벤치마크다.

huggingface.co
#token-efficiency#llm#semiconductors#applications
Enabling physician-centered oversight for AMIE
Article2025년 8월 12일

Enabling physician-centered oversight for AMIE

g AMIE는 환자 문진과 의료 기록 초안을 맡되 개별화된 의학적 조언은 금지하고, 최종 판단과 환자 전달은 감독 의사가 검토·수정하도록 설계된 AMIE의 의사 중심 감독 연구 프레임워크다.

research.google
#agent-routing#llm#semiconductors#applications
Implementing MCP Servers in Python: An AI Shopping Assistant with Gradio
Article2025년 7월 31일

Implementing MCP Servers in Python: An AI Shopping Assistant with Gradio

Gradio의 MCP 통합을 이용해 웹 탐색 도구와 IDM VTON 가상 피팅 모델을 연결하고, VS Code AI 채팅에서 사용할 수 있는 개인용 AI 쇼핑 도우미를 구현하는 방법을 설명한다.

huggingface.co
#llm#semiconductors#applications#agent-memory
Introducing AI Sheets: a tool to work with datasets using open AI models!
Article2025년 7월 31일

Introducing AI Sheets: a tool to work with datasets using open AI models!

Hugging Face AI Sheets는 스프레드시트형 인터페이스에서 프롬프트와 공개 AI 모델을 활용해 데이터셋을 만들고, 변환하고, 보강하고, 평가할 수 있게 해주는 오픈소스 노코드 도구다.

huggingface.co
#privacy-design#context-compression#prompt-library#llm
Lessons from Building an AI App Builder on Convex
Article2025년 7월 31일

Lessons from Building an AI App Builder on Convex

Chef의 사례는 AI 코딩 에이전트가 잘 작동하려면 좋은 추상화, 제한된 선택지, 타입 안전한 피드백 루프, 검증 가능한 평가 체계가 필요하다는 점을 보여준다.

stack.convex.dev
#ai-architecture#agent-routing#context-compression#prompt-library
Benchmarking Language Model Performance on 5th Gen Xeon at GCP
Article2025년 7월 29일

Benchmarking Language Model Performance on 5th Gen Xeon at GCP

Google Cloud의 5세대 Xeon 기반 C4 인스턴스는 3세대 Xeon 기반 N2보다 텍스트 임베딩과 텍스트 생성에서 큰 처리량 및 비용 효율 우위를 보이며, 경량 에이전틱 AI를 CPU만으로 배포할 가능성을 보여준다.

huggingface.co
#llm#semiconductors#applications#long-context
SensorLM: Learning the language of wearable sensors
Article2025년 7월 28일

SensorLM: Learning the language of wearable sensors

센서엘엠은 5,970만 시간의 웨어러블 센서 데이터와 자동 생성 설명문을 학습해 신체 신호를 자연어로 해석하고 검색·분류·설명하는 센서 언어 기반 모델군이다.

research.google
#ai-architecture#multimodal#llm#semiconductors
4M Models Scanned: Protect AI + Hugging Face 6 Months In
Article2025년 7월 27일

4M Models Scanned: Protect AI + Hugging Face 6 Months In

Protect AI와 Hugging Face의 6개월 파트너십은 Guardian 스캔으로 447만 개 모델 버전을 검사하고 35.2만 건의 unsafe/suspicious 이슈를 찾아, 공개 모델 사용자가 보안 위험을 더 명확히 판단하도록 돕고 있다.

huggingface.co
#service-design#llm#semiconductors#applications
TimeScope: How Long Can Your Video Large Multimodal Model Go?
Article2025년 7월 23일

TimeScope: How Long Can Your Video Large Multimodal Model Go?

TimeScope는 1분부터 8시간까지의 영상에 짧은 동영상 클립을 삽입해 검색·정보 종합·세밀한 시간 지각 능력을 측정하며, 최신 비전 언어 모델의 장시간 영상 이해가 아직 제한적임을 보여주는 오픈소스 벤치마크다.

huggingface.co
#multimodal#llm#semiconductors#vision-language-models
How Zapier Powers AI Chatbots with Web Knowledge Using Firecrawl
Article2025년 7월 21일

How Zapier Powers AI Chatbots with Web Knowledge Using Firecrawl

Zapier는 Firecrawl을 Zapier Chatbots에 통합해 고객의 공개 웹사이트와 헬프센터 콘텐츠를 코드 작성 없이 챗봇 지식으로 연결한다.

Eric Ciarla
#llm#semiconductors#applications#agent-deployment
Invideo AI enables anyone with an idea to produce high-quality videos
Article2025년 7월 17일

Invideo AI enables anyone with an idea to produce high-quality videos

인비디오 AI는 GPT‑4.1, gpt image 1, 텍스트 음성 변환 모델을 조율해 아이디어만으로 전문 품질의 영상을 빠르게 만들고 편집하게 하는 서비스다.

openai.com
#agent-routing#context-compression#prompt-library#search-advertising
Custom Kernels for All from Codex and Claude
Article2025년 7월 16일

Custom Kernels for All from Codex and Claude

CUDA 커널 개발 지식을 에이전트 스킬로 구조화해 Claude와 Codex가 실제 diffusers·transformers 대상의 커널, PyTorch 바인딩, 빌드 구성, 벤치마크까지 완성하도록 한 사례다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
Ettin Suite: SoTA Paired Encoders and Decoders
Article2025년 7월 16일

Ettin Suite: SoTA Paired Encoders and Decoders

에틴은 동일한 공개 데이터 2조 토큰, 모델 구조, 학습 절차를 적용한 1,700만~10억 매개변수 규모의 인코더·디코더 쌍으로, 두 아키텍처를 공정하게 비교하면서 양쪽 모두에서 공개 데이터 기반 최고 수준의 성능을 달성한 모델군이다.

huggingface.co
#ai-architecture#llm#semiconductors#applications
Migrating the Hub from Git LFS to Xet
Article2025년 7월 15일

Migrating the Hub from Git LFS to Xet

허깅페이스는 기존 사용자의 작업 방식을 유지하는 브리지와 무중단 백그라운드 마이그레이션을 기반으로, 50만 개 저장소와 20페타바이트 규모의 허브를 깃 대용량 파일 저장소에서 젯으로 전환하고 있다.

huggingface.co
#agent-routing#llm#semiconductors#applications
Building the Hugging Face MCP Server
Article2025년 7월 10일

Building the Hugging Face MCP Server

허깅페이스는 사용자가 도구와 Gradio 애플리케이션을 맞춤 구성할 수 있는 공식 MCP 서버를 구축하면서, 변화가 빠른 MCP 전송 규격 가운데 상태 비저장·직접 응답 방식의 Streamable HTTP를 운영 환경에 선택한 이유와 실제 배포 경험을 설명한다.

huggingface.co
#multimodal#agent-deployment#llm#semiconductors
Reachy Mini - The Open-Source Robot for Today's and Tomorrow's AI Builders
Article2025년 7월 9일

Reachy Mini - The Open-Source Robot for Today's and Tomorrow's AI Builders

Reachy Mini는 Pollen Robotics와 Hugging Face가 공개한 데스크톱 크기의 오픈소스 로봇으로, 인간 로봇 상호작용과 AI 실험을 Python 기반으로 쉽게 시도하도록 설계된 제품입니다.

huggingface.co
#privacy-design#multimodal#llm#semiconductors
Upskill your LLMs With Gradio MCP Servers
Article2025년 7월 9일

Upskill your LLMs With Gradio MCP Servers

Gradio의 MCP 지원을 활용하면 Hugging Face Spaces의 다양한 AI 도구를 Cursor 같은 LLM 클라이언트에 연결해 이미지 편집, 영상 전사, OCR, 음성 합성 등의 새로운 기능을 부여할 수 있습니다.

huggingface.co
#capex-cycle#llm#semiconductors#applications
Migrating data from Postgres to Convex
Article2025년 7월 8일

Migrating data from Postgres to Convex

Postgres 데이터를 Convex로 이전하는 방법은 소규모 데이터의 JSONL 덤프·가져오기에서 시작해, 스키마 정의와 관계 필드의 Convex ID 전환, 필요 시 Airbyte 기반 스트리밍 가져오기까지 이어진다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
Training and Finetuning Reranker Models with Sentence Transformers
Article2025년 7월 5일

Training and Finetuning Reranker Models with Sentence Transformers

이 글은 Sentence Transformers로 reranker 또는 Cross Encoder 모델을 도메인 데이터에 맞게 학습·파인튜닝하는 구성요소, 데이터 형식, hard negative mining의 중요성을 설명하고, 저자가 학습한 ModernBERT 기반 reranker가 자신의 평가 데이터에서 기존 공개 모델들을 앞섰다고 소개한다.

huggingface.co
#multimodal#capex-cycle#llm#vision-language-models
Exploring Quantization Backends in Diffusers
Article2025년 6월 27일

Exploring Quantization Backends in Diffusers

Diffusers는 Flux의 핵심 구성 요소를 여러 백엔드로 양자화해 이미지 품질을 크게 훼손하지 않으면서 메모리 사용량을 줄일 수 있으며, 정밀도와 백엔드에 따라 속도·메모리·사용 편의성의 차이가 뚜렷하다.

huggingface.co
#ai-architecture#multimodal#agent-memory#context-compression
Introducing Three New Serverless Inference Providers: Hyperbolic, Nebius AI Studio, and Novita 🔥
Article2025년 6월 27일

Introducing Three New Serverless Inference Providers: Hyperbolic, Nebius AI Studio, and Novita 🔥

Hugging Face Hub가 Hyperbolic, Nebius AI Studio, Novita를 새 서버리스 추론 제공자로 추가해 모델 페이지와 JS·Python SDK에서 여러 모델을 더 쉽게 선택해 사용할 수 있게 했다.

huggingface.co
#capex-cycle#llm#semiconductors#applications
이전1…12131415161714 / 17다음