Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#context-compression
Tag404건YouTube 25Article 379

#context-compression

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 286 · 연관도 60%#prompt-library공동문서 154 · 연관도 53%#semiconductors공동문서 354 · 연관도 50%#applications공동문서 334 · 연관도 50%#retrieval-index공동문서 160 · 연관도 49%#llm공동문서 342 · 연관도 47%#agent-routing공동문서 116 · 연관도 25%#agent-deployment공동문서 80 · 연관도 21%#ai-architecture공동문서 92 · 연관도 21%#compute공동문서 34 · 연관도 17%
Bringing Robotics AI to Embedded Platforms: Dataset Recording, VLA Fine‑Tuning, and On‑Device Optimizations
Article2024년 11월 28일

Bringing Robotics AI to Embedded Platforms: Dataset Recording, VLA Fine‑Tuning, and On‑Device Optimizations

NXP는 로봇 VLA를 임베디드 플랫폼에 배포하기 위해 신뢰도 높은 데이터셋 수집, ACT·SmolVLA 미세조정, i.MX 95에서의 모델 분해·양자화·비동기 추론 최적화가 함께 필요하다고 설명한다.

huggingface.co
#multimodal#agent-deployment#agent-memory#context-compression
Letting Large Models Debate: The First Multilingual LLM Debate Competition
Article2024년 11월 13일

Letting Large Models Debate: The First Multilingual LLM Debate Competition

BAAI는 정적 벤치마크와 사용자 투표형 아레나의 한계를 보완하기 위해 대형 언어모델들이 다국어로 직접 토론하며 추론력·논증력·언어 능력을 드러내는 FlagEval Debate를 제안한다.

huggingface.co
#anthropic#multimodal#context-compression#prompt-library
Launch Week II - Day 2: Introducing Location and Language Settings
Article2024년 10월 29일

Launch Week II - Day 2: Introducing Location and Language Settings

Firecrawl은 Launch Week II 2일차에 웹 스크래핑 결과를 국가와 선호 언어에 맞춰 받을 수 있는 Location and Language Settings 기능을 공개했다.

Eric Ciarla
#llm#semiconductors#applications#agent-deployment
Launch Week II - Day 1: Announcing the Batch Scrape Endpoint
Article2024년 10월 28일

Launch Week II - Day 1: Announcing the Batch Scrape Endpoint

Firecrawl은 Launch Week II 첫날 여러 URL을 한 번에 스크랩할 수 있는 Batch Scrape 엔드포인트를 공개하고, 동기·비동기 처리 방식과 사용 예시를 함께 안내했다.

Eric Ciarla
#agent-routing#llm#semiconductors#applications
Rearchitecting Hugging Face Uploads and Downloads
Article2024년 9월 27일

Rearchitecting Hugging Face Uploads and Downloads

허깅페이스는 대용량 모델·데이터셋 전송의 한계를 극복하기 위해 바이트 단위 중복 제거와 검증을 수행하는 콘텐츠 주소 기반 저장소(CAS)를 도입하고, 전 세계 업로드 분포에 맞춰 세 지역에 노드를 배치하는 새로운 업로드·다운로드 구조를 설계했다.

huggingface.co
#service-design#ai-architecture#llm#semiconductors
Bamba: Inference-Efficient Hybrid Mamba2 Model
Article2024년 9월 25일

Bamba: Inference-Efficient Hybrid Mamba2 Model

Bamba 9B는 KV cache 병목을 줄이기 위해 Mamba2와 트랜스포머를 결합한 9B급 하이브리드 모델로, 공개 데이터로 학습되고 주요 오픈소스 추론·학습 생태계에서 바로 실험할 수 있도록 공개된 모델이다.

huggingface.co
#model-scaling#ai-architecture#agent-memory#context-compression
Mini-R1: Reproduce Deepseek R1 „aha moment“ a RL tutorial
Article2024년 9월 25일

Mini-R1: Reproduce Deepseek R1 „aha moment“ a RL tutorial

3B 규모의 큐원 모델을 카운트다운 산술 과제와 그룹 상대 정책 최적화로 훈련해, 별도의 인간 피드백 없이 풀이를 검토하고 탐색하는 딥시크 R1의 작은 ‘아하 순간’을 재현한 강화학습 실험이다.

huggingface.co
#nvidia#service-design#agent-memory#capex-cycle
Expert Support case study: Bolstering a RAG app with LLM-as-a-Judge
Article2024년 9월 13일

Expert Support case study: Bolstering a RAG app with LLM-as-a-Judge

Digital Green은 농업 RAG 챗봇 Farmer.chat의 응답 품질을 대규모로 평가하기 위해 LLM as a Judge 체계를 구축하고, 충실도와 답변률의 균형을 기준으로 생성 모델을 비교·선정했다.

huggingface.co
#service-design#ai-architecture#agent-memory#context-compression
The Open Arabic LLM Leaderboard 2
Article2024년 8월 27일

The Open Arabic LLM Leaderboard 2

오픈 아랍어 LLM 리더보드 2는 기존 평가의 번역 편향·포화·검증 오류를 바로잡고, 아랍어 고유 특성과 실제 활용을 반영한 원어·인간 검수 벤치마크와 생성형 평가를 확대한 개편판이다.

huggingface.co
#agent-memory#context-compression#retrieval-index#llm
Announcing Fire Engine for Firecrawl
Article2024년 8월 6일

Announcing Fire Engine for Firecrawl

Firecrawl은 외부 스크래핑 서비스의 실패와 속도 문제를 해결하기 위해 자체 기본 백엔드인 Fire Engine을 도입했다고 발표했다.

Eric Ciarla
#llm#semiconductors#applications#agent-deployment
Firecrawl June 2024 Updates
Article2024년 6월 30일

Firecrawl June 2024 Updates

Firecrawl의 2024년 6월 업데이트는 Gamma, Dify, Praison, Flowise 연동, 새 대시보드, 플랫폼 기능 개선, 신규 튜토리얼 공개를 중심으로 진행됐다.

Nicolas Camara
#agent-routing#llm#semiconductors#applications
How much does it cost to train frontier AI models?
Article2024년 6월 3일

How much does it cost to train frontier AI models?

Epoch AI는 프런티어 AI 모델의 최종 훈련 비용이 2016년 이후 연 2~3배 수준으로 빠르게 증가해 왔으며, 현재 추세가 이어지면 2027년 최대 규모 훈련 실행은 10억 달러를 넘을 수 있다고 분석한다.

Ben Cottier
#llm#semiconductors#applications#agent-deployment
Do the returns to software R&D point towards a singularity?
Article2024년 5월 17일

Do the returns to software R&D point towards a singularity?

이 글은 소프트웨어 연구개발의 수익률이 1을 넘을 경우 AI가 자기 연구개발을 자동화하면서 초쌍곡적 성장, 즉 소프트웨어 특이점으로 이어질 수 있는지 이론과 제한적 실증 자료로 검토한다.

Tamay Besiroglu
#luxury-hospitality#luxury-travel#travel-hospitality#llm
Introducing SyGra Studio
Article2024년 4월 4일

Introducing SyGra Studio

SyGra Studio는 합성 데이터 생성 흐름을 시각적으로 구성하고, 생성되는 설정과 실행 상태·비용·결과를 한 화면에서 확인할 수 있게 만든 대화형 작업 환경이다.

huggingface.co
#ai-architecture#agent-routing#context-compression#prompt-library
🐯 Liger GRPO meets TRL
Article2024년 2월 5일

🐯 Liger GRPO meets TRL

라이거의 청크 단위 손실 계산을 TRL의 그룹 상대 정책 최적화에 통합해 모델 품질을 유지하면서 최대 GPU 메모리를 40% 줄이고, 분산 학습과 매개변수 효율적 미세 조정 및 고속 생성 서버 연계까지 지원한 결과를 소개한다.

huggingface.co
#agent-deployment#agent-memory#context-compression#retrieval-index
All About the U.S. Executive Order on AI Use and Development
Article2023년 11월 1일

All About the U.S. Executive Order on AI Use and Development

바이든 행정부의 AI 행정명령은 국방·비상권한에 근거해 일부 기초 모델의 보고·시험을 요구하고, 연방기관이 안전·프라이버시·차별·경쟁력·국제표준 관련 기준을 마련하도록 지시한 조치다.

@DeepLearningAI
#anthropic#privacy-design#llm#semiconductors
Explosive growth from AI: A review of the arguments
Article2023년 9월 23일

Explosive growth from AI: A review of the arguments

Epoch AI 글은 고도화된 AI가 인간 노동을 대체할 때 폭발적 경제성장이 가능하다는 성장모형의 논거와, 규제·개발 난도·정렬 문제 같은 반론을 함께 검토하며 가능성은 배제하기 어렵지만 확실하지도 않다고 정리한다.

Ege Erdil
#llm#semiconductors#applications#agent-deployment
AI Canon
Article2023년 5월 25일

AI Canon

a16z의 ‘AI Canon’은 현대 AI를 이해하기 위해 영향력이 컸던 논문, 글, 강의, 실무 가이드, 시장 분석 자료를 단계별로 묶은 선별 참고 목록이다.

Derrick Harris, Matt Bornstein, Guido Appenzeller
#anthropic#privacy-design#ai-architecture#agent-routing
Will we run out of ML data? Projecting dataset size trends
Article2022년 11월 10일

Will we run out of ML data? Projecting dataset size trends

Epoch AI는 현재의 ML 데이터 사용 증가 추세가 지속되고 데이터 효율성 혁신이 없다는 가정 아래, 고품질 언어 데이터는 2026년 전, 저품질 언어 데이터와 이미지 데이터는 2030~2060년 사이에 병목이 될 수 있다고 전망한다.

Pablo Villalobos
#llm#semiconductors#applications#agent-memory
TRL v1.0: Post-Training Library Built to Move with the Field
Article2017년 7월 20일

TRL v1.0: Post-Training Library Built to Move with the Field

TRL v1.0은 빠르게 변하는 포스트트레이닝 분야에서 안정적인 핵심 API와 실험적 방법을 함께 수용하도록 설계된 Hugging Face 생태계의 범용 포스트트레이닝 라이브러리다.

huggingface.co
#service-design#ai-architecture#agent-memory#context-compression
이전1…15161717 / 17다음