Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#retrieval-index
Tag248건YouTube 17Article 231

#retrieval-index

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 248 · 연관도 66%#context-compression공동문서 160 · 연관도 50%#applications공동문서 197 · 연관도 38%#llm공동문서 213 · 연관도 37%#semiconductors공동문서 203 · 연관도 37%#agent-routing공동문서 103 · 연관도 28%#ai-architecture공동문서 82 · 연관도 24%#service-design공동문서 41 · 연관도 14%#gpu공동문서 17 · 연관도 12%#multimodal공동문서 25 · 연관도 11%
Parquet Content-Defined Chunking
Article2025년 2월 19일

Parquet Content-Defined Chunking

파케이 콘텐츠 정의 청킹은 유사한 파케이 파일에서 변경된 데이터 조각만 전송하도록 해, 허깅 페이스 허브의 업로드·다운로드 시간과 저장 비용을 크게 줄이는 기능이다.

huggingface.co
#agent-routing#semiconductors#applications#compute
Types and Validators: A Convex Cookbook
Article2025년 2월 15일

Types and Validators: A Convex Cookbook

이 글은 Convex에서 스키마 validator, 함수 인자 검증, TypeScript 타입을 중복 없이 연결해 데이터 일관성과 개발 편의성을 높이는 방법을 요리책 예시로 설명한다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
AI and the Future of Cybersecurity: Why Openness Matters
Article2025년 2월 4일

AI and the Future of Cybersecurity: Why Openness Matters

이 글은 Mythos 사례를 통해 AI 사이버보안의 핵심이 단일 모델이 아니라 시스템·생태계에 있으며, 방어자가 공격자와 맞서기 위해서는 개방형 도구와 감사 가능한 구조가 중요하다고 설명한다.

huggingface.co
#llm#semiconductors#ai-coding#long-context
DABStep: Data Agent Benchmark for Multi-step Reasoning
Article2025년 2월 4일

DABStep: Data Agent Benchmark for Multi-step Reasoning

DABstep은 실제 결제 데이터 분석 업무에서 나온 450개 이상의 과제를 통해, 현재 LLM 기반 데이터 에이전트가 다단계 추론·도메인 이해·코드 실행을 결합한 현실적 분석 문제를 얼마나 해결할 수 있는지 평가하는 벤치마크입니다.

huggingface.co
#multimodal#agent-routing#llm#semiconductors
Hugging Face and JFrog partner to make AI Security more transparent
Article2025년 2월 4일

Hugging Face and JFrog partner to make AI Security more transparent

Hugging Face는 JFrog 스캐너를 Hub에 통합해 모델 파일 속 코드의 악성 사용 가능성을 더 깊게 분석하고, ML 커뮤니티의 모델 공유 보안을 강화한다고 발표했다.

huggingface.co
#agent-deployment#llm#semiconductors#applications
Why Convex Queries are the Ultimate Form of Derived State
Article2025년 1월 27일

Why Convex Queries are the Ultimate Form of Derived State

이 글은 파생 상태를 로컬 React 계산에서 서버 기반 다중 클라이언트 동기화 문제로 확장해 설명하고, Convex 쿼리가 반응형 데이터베이스 위에서 프런트엔드가 필요한 파생 상태를 항상 최신으로 제공하는 방식이라고 주장한다.

stack.convex.dev
#agent-routing#semiconductors#applications#compute
We now support VLMs in smolagents!
Article2025년 1월 24일

We now support VLMs in smolagents!

smolagents에 비전 지원이 추가되어 VLM을 에이전트 파이프라인에서 기본적으로 활용하고, 특히 웹 브라우징처럼 시각 정보가 중요한 작업을 수행할 수 있게 되었습니다.

huggingface.co
#multimodal#agent-memory#context-compression#retrieval-index
Are better models better? — Benedict Evans
Article2025년 1월 22일

Are better models better? — Benedict Evans

이 글은 더 좋은 생성형 AI 모델이 항상 더 유용한 것은 아니며, 특히 정답이 하나인 결정론적 업무에서는 ‘더 그럴듯한 답’이 아니라 검증 가능한 ‘맞는 답’이 필요하다고 설명합니다.

ben-evans.com
#agent-routing#llm#semiconductors#applications
Timm ❤️ Transformers: Use any timm model with transformers
Article2025년 1월 15일

Timm ❤️ Transformers: Use any timm model with transformers

TimmWrapper는 방대한 timm 비전 모델을 Transformers의 파이프라인, 자동 클래스, 양자화, Trainer, LoRA 워크플로에서 사용하고 다시 timm으로 불러올 수 있게 연결합니다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
Open R1: How to use OlympicCoder locally for coding
Article2025년 1월 12일

Open R1: How to use OlympicCoder locally for coding

올림픽코더 7B의 4비트 양자화 모델을 엘엠 스튜디오에서 구동하고 컨티뉴 확장 기능으로 비주얼 스튜디오 코드에 연결해 로컬 코딩 도우미로 사용하는 방법을 설명한다.

huggingface.co
#llm#semiconductors#applications#long-context
CO₂ Emissions and Models Performance: Insights from the Open LLM Leaderboard
Article2025년 1월 9일

CO₂ Emissions and Models Performance: Insights from the Open LLM Leaderboard

Hugging Face Open LLM Leaderboard의 평가 데이터를 통해 모델 추론 시 CO₂ 배출은 대체로 모델 크기와 함께 늘지만, 성능 향상은 그만큼 비례하지 않으며 커뮤니티 파인튜닝 모델이 더 효율적인 경우가 많다는 점을 분석했다.

huggingface.co
#hotel-review#privacy-design#travel-hospitality#llm
State of open video generation models in Diffusers
Article2025년 1월 8일

State of open video generation models in Diffusers

오픈 비디오 생성 모델은 빠르게 발전하고 있지만 높은 메모리 요구량, 긴 생성 시간, 제한적인 일반화가 확산을 가로막고 있으며, 디퓨저스는 양자화·오프로딩·분할 추론을 조합해 이러한 모델을 더 적은 자원에서 실행할 수 있도록 지원한다.

huggingface.co
#ai-architecture#agent-memory#context-compression#retrieval-index
Introducing smolagents: simple agents that write actions in code.
Article2024년 12월 27일

Introducing smolagents: simple agents that write actions in code.

smolagents는 언어 모델이 외부 도구를 코드로 조합·실행하도록 지원해, 복잡한 다단계 에이전트 워크플로를 단순하게 구축하는 라이브러리다.

huggingface.co
#anthropic#agent-routing#travel-hospitality#llm
Mastering Long Contexts in LLMs with KVPress
Article2024년 12월 21일

Mastering Long Contexts in LLMs with KVPress

KVPress는 긴 문맥에서 선형으로 증가하는 키·값 캐시를 중요도 기반으로 압축해 메모리 사용량을 줄이고 생성 속도를 높이는 모듈형 도구다.

huggingface.co
#ai-architecture#multimodal#agent-memory#context-compression
Add a collaborative document editor to your app
Article2024년 12월 19일

Add a collaborative document editor to your app

단순한 <textarea 로는 동시 편집이 충돌하므로, Convex와 ProseMirror 기반 동기화 컴포넌트, Tiptap 또는 BlockNote를 조합해 기존 앱에 협업 문서 편집기를 붙이는 절차를 설명한다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
Fast LoRA inference for Flux with Diffusers and PEFT
Article2024년 12월 5일

Fast LoRA inference for Flux with Diffusers and PEFT

Diffusers와 PEFT의 LoRA 핫스와핑, torch.compile, FP8 양자화, Flash Attention 3를 조합해 Flux.1 Dev의 재컴파일 문제를 피하고 고성능 GPU에서 약 2.23배, RTX 4090에서 약 2.04배 빠른 LoRA 추론을 구현한 최적화 방법을 설명한다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs
Article2024년 12월 5일

How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs

간단한 일정 관리 대화 실험에서 대규모 언어 모델의 실수 수정 능력을 비교한 결과, 최신 3B~9B급 모델은 대체로 피드백을 반영했지만 소형·구형 모델은 형식 준수와 문맥 유지에서 반복적으로 실패했으며 모델 크기만으로 성능이 보장되지는 않았다.

huggingface.co
#agent-memory#context-compression#retrieval-index#llm
Bringing Robotics AI to Embedded Platforms: Dataset Recording, VLA Fine‑Tuning, and On‑Device Optimizations
Article2024년 11월 28일

Bringing Robotics AI to Embedded Platforms: Dataset Recording, VLA Fine‑Tuning, and On‑Device Optimizations

NXP는 로봇 VLA를 임베디드 플랫폼에 배포하기 위해 신뢰도 높은 데이터셋 수집, ACT·SmolVLA 미세조정, i.MX 95에서의 모델 분해·양자화·비동기 추론 최적화가 함께 필요하다고 설명한다.

huggingface.co
#multimodal#agent-deployment#agent-memory#context-compression
Going local-first with Automerge and Convex
Article2024년 11월 19일

Going local-first with Automerge and Convex

이 글은 Automerge의 CRDT 기반 로컬 우선 편집 모델과 Convex 백엔드를 함께 사용해 오프라인 편집, 충돌 없는 협업, 서버 기반 권한·관계·동기화를 결합하는 방법을 설명한다.

stack.convex.dev
#ai-architecture#agent-routing#llm#semiconductors
An Object Sync Engine for Local-first Apps
Article2024년 11월 13일

An Object Sync Engine for Local-first Apps

이 글은 로컬 퍼스트 웹 앱에 적합한 객체 동기화 엔진을 정의하고, 로컬 저장소·서버 저장소·동기화 프로토콜의 역할을 기존 사례와 컨벡스의 방향으로 비교한다.

stack.convex.dev
#service-design#ai-architecture#agent-memory#retrieval-index
Launch Week II - Day 2: Introducing Location and Language Settings
Article2024년 10월 29일

Launch Week II - Day 2: Introducing Location and Language Settings

Firecrawl은 Launch Week II 2일차에 웹 스크래핑 결과를 국가와 선호 언어에 맞춰 받을 수 있는 Location and Language Settings 기능을 공개했다.

Eric Ciarla
#llm#semiconductors#applications#agent-deployment
Launch Week II - Day 1: Announcing the Batch Scrape Endpoint
Article2024년 10월 28일

Launch Week II - Day 1: Announcing the Batch Scrape Endpoint

Firecrawl은 Launch Week II 첫날 여러 URL을 한 번에 스크랩할 수 있는 Batch Scrape 엔드포인트를 공개하고, 동기·비동기 처리 방식과 사용 예시를 함께 안내했다.

Eric Ciarla
#agent-routing#llm#semiconductors#applications
Bamba: Inference-Efficient Hybrid Mamba2 Model
Article2024년 9월 25일

Bamba: Inference-Efficient Hybrid Mamba2 Model

Bamba 9B는 KV cache 병목을 줄이기 위해 Mamba2와 트랜스포머를 결합한 9B급 하이브리드 모델로, 공개 데이터로 학습되고 주요 오픈소스 추론·학습 생태계에서 바로 실험할 수 있도록 공개된 모델이다.

huggingface.co
#model-scaling#ai-architecture#agent-memory#context-compression
Mini-R1: Reproduce Deepseek R1 „aha moment“ a RL tutorial
Article2024년 9월 25일

Mini-R1: Reproduce Deepseek R1 „aha moment“ a RL tutorial

3B 규모의 큐원 모델을 카운트다운 산술 과제와 그룹 상대 정책 최적화로 훈련해, 별도의 인간 피드백 없이 풀이를 검토하고 탐색하는 딥시크 R1의 작은 ‘아하 순간’을 재현한 강화학습 실험이다.

huggingface.co
#nvidia#service-design#agent-memory#capex-cycle
이전1…89101110 / 11다음