Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#context-compression
Tag404건YouTube 25Article 379

#context-compression

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 286 · 연관도 60%#prompt-library공동문서 154 · 연관도 53%#semiconductors공동문서 354 · 연관도 50%#applications공동문서 334 · 연관도 50%#retrieval-index공동문서 160 · 연관도 49%#llm공동문서 342 · 연관도 47%#agent-routing공동문서 116 · 연관도 25%#agent-deployment공동문서 80 · 연관도 21%#ai-architecture공동문서 92 · 연관도 21%#compute공동문서 34 · 연관도 17%
Translate SQL into Convex Queries
Article2025년 3월 19일

Translate SQL into Convex Queries

이 글은 SQL에 익숙한 개발자가 UNION, WHERE 필터, JOIN, DISTINCT, GROUP BY 같은 패턴을 Convex 쿼리와 QueryStreams 방식으로 옮기는 방법을 Slack형 채팅 앱 예시로 설명한다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
Trace & Evaluate your Agent with Arize Phoenix
Article2025년 3월 1일

Trace & Evaluate your Agent with Arize Phoenix

에이전트의 내부 실행 과정을 추적하고 도구 호출 결과를 체계적으로 평가해, 단순히 작동하는 수준을 넘어 실제로 효과적인 에이전트로 개선하는 방법을 아라이즈 피닉스와 스몰에이전츠 예제로 설명한다.

huggingface.co
#agent-routing#api-gpt-4o#llm#semiconductors
Visualize and understand GPU memory in PyTorch
Article2025년 3월 1일

Visualize and understand GPU memory in PyTorch

PyTorch 메모리 스냅샷으로 GPU 사용량을 단계별로 시각화하고, 학습 과정의 최대 메모리를 구성 요소별로 추정하는 방법을 설명한다.

huggingface.co
#ai-architecture#agent-memory#capex-cycle#context-compression
HuggingFace, IISc partner to supercharge model building on India's diverse languages
Article2025년 2월 27일

HuggingFace, IISc partner to supercharge model building on India's diverse languages

허깅페이스와 인도과학원·아트파크는 인도 전역의 언어·방언·지역·인구통계적 다양성을 담은 오픈소스 다중양식 데이터셋 ‘바니’의 접근성과 활용성을 높이기 위해 협력한다.

huggingface.co
#multimodal#llm#semiconductors#vision-language-models
Parquet Content-Defined Chunking
Article2025년 2월 19일

Parquet Content-Defined Chunking

파케이 콘텐츠 정의 청킹은 유사한 파케이 파일에서 변경된 데이터 조각만 전송하도록 해, 허깅 페이스 허브의 업로드·다운로드 시간과 저장 비용을 크게 줄이는 기능이다.

huggingface.co
#agent-routing#semiconductors#applications#compute
Types and Validators: A Convex Cookbook
Article2025년 2월 15일

Types and Validators: A Convex Cookbook

이 글은 Convex에서 스키마 validator, 함수 인자 검증, TypeScript 타입을 중복 없이 연결해 데이터 일관성과 개발 편의성을 높이는 방법을 요리책 예시로 설명한다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
AI and the Future of Cybersecurity: Why Openness Matters
Article2025년 2월 4일

AI and the Future of Cybersecurity: Why Openness Matters

이 글은 Mythos 사례를 통해 AI 사이버보안의 핵심이 단일 모델이 아니라 시스템·생태계에 있으며, 방어자가 공격자와 맞서기 위해서는 개방형 도구와 감사 가능한 구조가 중요하다고 설명한다.

huggingface.co
#llm#semiconductors#ai-coding#long-context
Hugging Face and JFrog partner to make AI Security more transparent
Article2025년 2월 4일

Hugging Face and JFrog partner to make AI Security more transparent

Hugging Face는 JFrog 스캐너를 Hub에 통합해 모델 파일 속 코드의 악성 사용 가능성을 더 깊게 분석하고, ML 커뮤니티의 모델 공유 보안을 강화한다고 발표했다.

huggingface.co
#agent-deployment#llm#semiconductors#applications
Why Convex Queries are the Ultimate Form of Derived State
Article2025년 1월 27일

Why Convex Queries are the Ultimate Form of Derived State

이 글은 파생 상태를 로컬 React 계산에서 서버 기반 다중 클라이언트 동기화 문제로 확장해 설명하고, Convex 쿼리가 반응형 데이터베이스 위에서 프런트엔드가 필요한 파생 상태를 항상 최신으로 제공하는 방식이라고 주장한다.

stack.convex.dev
#agent-routing#semiconductors#applications#compute
We now support VLMs in smolagents!
Article2025년 1월 24일

We now support VLMs in smolagents!

smolagents에 비전 지원이 추가되어 VLM을 에이전트 파이프라인에서 기본적으로 활용하고, 특히 웹 브라우징처럼 시각 정보가 중요한 작업을 수행할 수 있게 되었습니다.

huggingface.co
#multimodal#agent-memory#context-compression#retrieval-index
Operator System Card
Article2025년 1월 23일

Operator System Card

오퍼레이터는 화면을 보고 브라우저를 조작하는 컴퓨터 사용 에이전트로, 오픈AI는 실제 행동에서 발생할 수 있는 피해를 줄이기 위해 위험 작업 거부, 중요 행동 전 사용자 확인, 외부 레드팀, 프런티어 위험 평가를 결합한 다층 안전 체계를 적용했다.

openai.com
#openai#privacy-design#multimodal#context-compression
Optimize Transaction Throughput: 3 Patterns for Scaling with Convex and ACID Databases
Article2025년 1월 23일

Optimize Transaction Throughput: 3 Patterns for Scaling with Convex and ACID Databases

이 글은 직렬화 가능한 트랜잭션의 충돌을 줄여 처리량을 높이는 방법으로 큐, hot/cold 테이블 분리, predicate locking 세 가지 패턴을 설명한다.

stack.convex.dev
#agent-routing#workflow-automation#semiconductors#applications
Are better models better? — Benedict Evans
Article2025년 1월 22일

Are better models better? — Benedict Evans

이 글은 더 좋은 생성형 AI 모델이 항상 더 유용한 것은 아니며, 특히 정답이 하나인 결정론적 업무에서는 ‘더 그럴듯한 답’이 아니라 검증 가능한 ‘맞는 답’이 필요하다고 설명합니다.

ben-evans.com
#agent-routing#llm#semiconductors#applications
Timm ❤️ Transformers: Use any timm model with transformers
Article2025년 1월 15일

Timm ❤️ Transformers: Use any timm model with transformers

TimmWrapper는 방대한 timm 비전 모델을 Transformers의 파이프라인, 자동 클래스, 양자화, Trainer, LoRA 워크플로에서 사용하고 다시 timm으로 불러올 수 있게 연결합니다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
Open R1: How to use OlympicCoder locally for coding
Article2025년 1월 12일

Open R1: How to use OlympicCoder locally for coding

올림픽코더 7B의 4비트 양자화 모델을 엘엠 스튜디오에서 구동하고 컨티뉴 확장 기능으로 비주얼 스튜디오 코드에 연결해 로컬 코딩 도우미로 사용하는 방법을 설명한다.

huggingface.co
#llm#semiconductors#applications#long-context
Tiny Agents in Python: a MCP-powered agent in ~70 lines of code
Article2025년 1월 12일

Tiny Agents in Python: a MCP-powered agent in ~70 lines of code

이 글은 Hugging Face의 huggingface hub에 포함된 MCP 클라이언트를 활용해 Python에서 약 70줄짜리 Tiny Agent를 실행하고 구성하며, LLM이 MCP 서버의 도구를 발견·호출·반영하는 흐름을 설명한다.

huggingface.co
#context-compression#prompt-library#llm#semiconductors
State of open video generation models in Diffusers
Article2025년 1월 8일

State of open video generation models in Diffusers

오픈 비디오 생성 모델은 빠르게 발전하고 있지만 높은 메모리 요구량, 긴 생성 시간, 제한적인 일반화가 확산을 가로막고 있으며, 디퓨저스는 양자화·오프로딩·분할 추론을 조합해 이러한 모델을 더 적은 자원에서 실행할 수 있도록 지원한다.

huggingface.co
#ai-architecture#agent-memory#context-compression#retrieval-index
Controlling Language Model Generation with NVIDIA's LogitsProcessorZoo
Article2024년 12월 23일

Controlling Language Model Generation with NVIDIA's LogitsProcessorZoo

이 글은 언어 모델의 다음 토큰 선택 과정에서 로짓을 직접 조정해 길이, 문맥 유지, 필수 문구, 선택형 답변 같은 생성 제약을 제어하는 방법을 NVIDIA의 LogitsProcessorZoo와 Hugging Face 생성 API 예시로 설명한다.

huggingface.co
#nvidia#privacy-design#agent-routing#capex-cycle
Mastering Long Contexts in LLMs with KVPress
Article2024년 12월 21일

Mastering Long Contexts in LLMs with KVPress

KVPress는 긴 문맥에서 선형으로 증가하는 키·값 캐시를 중요도 기반으로 압축해 메모리 사용량을 줄이고 생성 속도를 높이는 모듈형 도구다.

huggingface.co
#ai-architecture#multimodal#agent-memory#context-compression
Add a collaborative document editor to your app
Article2024년 12월 19일

Add a collaborative document editor to your app

단순한 <textarea 로는 동시 편집이 충돌하므로, Convex와 ProseMirror 기반 동기화 컴포넌트, Tiptap 또는 BlockNote를 조합해 기존 앱에 협업 문서 편집기를 붙이는 절차를 설명한다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
Welcome to the Falcon 3 Family of Open Models!
Article2024년 12월 17일

Welcome to the Falcon 3 Family of Open Models!

팰컨3는 10억~100억 매개변수 규모에서 과학·수학·코딩·추론 성능과 학습 효율을 함께 높이고, 다양한 배포 형식과 개방형 라이선스를 제공하는 디코더 전용 언어 모델 제품군이다.

huggingface.co
#ai-architecture#llm#semiconductors#applications
Hugging Face models in Amazon Bedrock
Article2024년 12월 9일

Hugging Face models in Amazon Bedrock

Models mentioned in this article 1 More Articles from our Blog aws partnerships How to deploy and fine tune DeepSeek models on AWS 56 January 30, 2025…

huggingface.co
#service-design#agent-deployment#llm#semiconductors
Fast LoRA inference for Flux with Diffusers and PEFT
Article2024년 12월 5일

Fast LoRA inference for Flux with Diffusers and PEFT

Diffusers와 PEFT의 LoRA 핫스와핑, torch.compile, FP8 양자화, Flash Attention 3를 조합해 Flux.1 Dev의 재컴파일 문제를 피하고 고성능 GPU에서 약 2.23배, RTX 4090에서 약 2.04배 빠른 LoRA 추론을 구현한 최적화 방법을 설명한다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs
Article2024년 12월 5일

How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs

간단한 일정 관리 대화 실험에서 대규모 언어 모델의 실수 수정 능력을 비교한 결과, 최신 3B~9B급 모델은 대체로 피드백을 반영했지만 소형·구형 모델은 형식 준수와 문맥 유지에서 반복적으로 실패했으며 모델 크기만으로 성능이 보장되지는 않았다.

huggingface.co
#agent-memory#context-compression#retrieval-index#llm
이전1…1415161716 / 17다음