Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#retrieval-index
Tag248건YouTube 17Article 231

#retrieval-index

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 248 · 연관도 66%#context-compression공동문서 160 · 연관도 50%#applications공동문서 197 · 연관도 38%#llm공동문서 213 · 연관도 37%#semiconductors공동문서 203 · 연관도 37%#agent-routing공동문서 103 · 연관도 28%#ai-architecture공동문서 82 · 연관도 24%#service-design공동문서 41 · 연관도 14%#gpu공동문서 17 · 연관도 12%#multimodal공동문서 25 · 연관도 11%
Profiling in PyTorch (Part 2): From nn.Linear to a Fused MLP
Article2026년 6월 11일

Profiling in PyTorch (Part 2): From nn.Linear to a Fused MLP

이 글은 프로파일러 추적을 통해 nn.Linear의 전치·행렬곱·편향 처리가 실제로 어떻게 실행되는지 분석하고, 단일 선형 계층에서 GeGLU MLP의 다중 커널 구조로 관찰 범위를 확장한다.

huggingface.co
#ai-architecture#agent-memory#capex-cycle#context-compression
MORE Hermes Agent + HyperFrames: Free AI Tools to Make Amazing Videos
YouTube2026년 6월 11일

MORE Hermes Agent + HyperFrames: Free AI Tools to Make Amazing Videos

Hermes Agent + HyperFrames는 무료 AI 영상 제작 도구로 꽤 실용적인 결과를 만들 수 있지만, 핵심은 “비슷한 효과”가 아니라 실제 HyperFrames catalog transition과 block을 정확히 호출하게 만드는 데 있다.

Tonbi''s AI Garage
#agent-memory#context-compression#retrieval-index#agent-systems
How frontier teams are reinventing AI-native development
Article2026년 6월 10일

How frontier teams are reinventing AI-native development

프런티어 팀은 AI를 단순 코딩 보조 도구가 아니라 소프트웨어 개발 방식의 기반으로 삼아, 에이전트가 잘 판단할 수 있는 맥락과 워크플로를 재설계함으로써 생산 배포 속도를 크게 높이고 있다.

aws.amazon.com
#agent-deployment#agent-routing#llm#semiconductors
A Local ChatGPT? Testing PewDiePie''s Odysseus (Setup Guide + Feature Tour)
YouTube2026년 6월 10일

A Local ChatGPT? Testing PewDiePie''s Odysseus (Setup Guide + Feature Tour)

Odysseus는 Local ChatGPT를 로컬 우선·프라이버시 중심으로 구현하려는 흥미로운 셀프호스팅 AI 작업공간이지만, Docker GPU 인식·모델 서빙·초기 보안 설정까지 직접 확인해야 실사용성이 드러난다.

Tonbi''s AI Garage
#anthropic-model-roadmap#frontier-model-evaluation#core-thesis#explainer
Better decisions at scale: How mathematical optimization delivers where intuition fails
Article2026년 6월 8일

Better decisions at scale: How mathematical optimization delivers where intuition fails

수학적 최적화는 직관과 단순 규칙으로 처리하기 어려운 대규모 운영 의사결정을 제약 조건 안에서 검증 가능한 최적 해로 바꾸는 AI 접근법이다.

aws.amazon.com
#agent-routing#llm#semiconductors#applications
DeepSeek-V4: a million-token context that agents can actually use
Article2026년 6월 8일

DeepSeek-V4: a million-token context that agents can actually use

DeepSeek V4는 최고 벤치마크 점수보다 100만 토큰 문맥을 실제 에이전트 작업에서 감당하게 만드는 긴 문맥 효율, 도구 호출 지속성, 샌드박스 기반 학습 인프라에 초점을 둔 모델이다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
How to Build a Multi-Agent Workflow for LLM Wikis in Hermes Kanban
YouTube2026년 6월 8일

How to Build a Multi-Agent Workflow for LLM Wikis in Hermes Kanban

Hermes Kanban 기반 Multi Agent Workflow는 LLM 위키를 더 최신이고 검증 가능한 지식베이스로 유지하기 위한 실전형 자동화 구조다.

Tonbi''s AI Garage
#anthropic-model-roadmap#frontier-model-evaluation#agent-systems#core-thesis
Designing the hf CLI as an agent-optimized way to work with the Hub
Article2026년 6월 4일

Designing the hf CLI as an agent-optimized way to work with the Hub

Hugging Face는 hf CLI를 사람과 코딩 에이전트가 모두 효율적으로 Hub를 다룰 수 있도록 재설계했고, 특히 복잡한 다단계 작업에서 에이전트의 토큰 사용과 실패를 줄인다고 설명한다.

huggingface.co
#agent-routing#llm#semiconductors#applications
Stanford Robotics Seminar ENGR319
YouTube2026년 6월 4일

Stanford Robotics Seminar ENGR319

Geometry in Robot Learning의 핵심은 로봇 학습을 무작정 더 큰 데이터와 모델로 밀어붙이기보다, 기하·대칭성·좌표계 구조를 모델에 넣어 데이터 효율성과 자세 일반화를 높이려는 것이다.

Stanford Online
#anthropic-model-roadmap#frontier-model-evaluation#agent-systems#core-thesis
Figma Exec on Why the SaaSpocalypse Is a Goldmine
Article2026년 6월 3일

Figma Exec on Why the SaaSpocalypse Is a Goldmine

Figma의 Matt Colyer는 AI가 SaaS를 무너뜨리기보다 개발자와 소프트웨어 수요를 폭발적으로 늘려 기존 제품에 더 큰 기회를 만든다고 본다.

Dan Shipper
#privacy-design#agent-routing#llm#semiconductors
Reading Today’s Headlines Through AI: A Real-Time Audit of Six Commercial Chatbots
Article2026년 6월 3일

Reading Today’s Headlines Through AI: A Real-Time Audit of Six Commercial Chatbots

스탠퍼드 HAI 연구는 상용 AI 챗봇이 당일 뉴스 질문에서 높은 평균 정확도를 보였지만, 실제 신뢰성은 언어·지역별 검색 인프라, 출처 선택, 불완전한 질문에 대한 취약성에 크게 좌우된다고 분석했다.

hai.stanford.edu
#privacy-design#llm#semiconductors#applications
‘AI gravity’ is pulling you toward dependency. Here’s how to push back
Article2026년 6월 2일

‘AI gravity’ is pulling you toward dependency. Here’s how to push back

MIT Sloan의 에릭 소는 AI가 업무 효율을 높이는 동시에 인간의 사고·학습·조직 지식을 약화시키는 ‘AI gravity’를 만들 수 있다고 경고하며, 기업과 개인이 의도적으로 인지 역량을 보존해야 한다고 말한다.

mitsloan.mit.edu
#agent-routing#llm#semiconductors#applications
Heeding the pope’s call to ensure AI protects human dignity
Article2026년 6월 1일

Heeding the pope’s call to ensure AI protects human dignity

MIT Sloan의 Thomas A. Kochan은 AI 시대에 인간 존엄을 지키려면 노동자가 AI 도입과 활용 방향을 결정하는 과정에 실질적으로 참여하고, 기술이 만든 경제적 이익을 함께 나누는 새로운 사회계약이 필요하다고 주장한다.

mitsloan.mit.edu
#semiconductors#applications#compute#agent-memory
April 2026: LangChain Newsletter
Article2026년 5월 28일

April 2026: LangChain Newsletter

2026년 4월 LangChain 뉴스레터는 LangSmith 평가·비용·도구 관리 업데이트, Deep Agents 배포 기능, Interrupt 2026 행사, 에이전트 개선 루프 콘텐츠와 고객 사례를 소개한다.

langchain.com
#multimodal#agent-deployment#agent-memory#context-compression
Readable TypeScript code: 14 patterns for humans and AI
Article2026년 5월 28일

Readable TypeScript code: 14 patterns for humans and AI

이 글은 인공지능 보조 도구가 빠르게 만든 타입스크립트 코드일수록 다음 사람이 읽고 고치기 쉬운 구조, 흐름, 타입 모델링이 더 중요하다고 설명한다.

stack.convex.dev
#agent-routing#llm#semiconductors#ai-coding
AI Hiring Tools Can Yield Racial Bias and Systemic Rejection
Article2026년 5월 26일

AI Hiring Tools Can Yield Racial Bias and Systemic Rejection

스탠퍼드 HAI가 실제 채용 데이터를 대규모로 분석한 결과, 인공지능 채용 선별 도구가 직무별 인종적 불균형을 만들고 같은 지원자를 여러 채용 과정에서 반복적으로 배제할 수 있다는 우려가 제기됐다.

hai.stanford.edu
#privacy-design#llm#semiconductors#ai-coding
Firecrawl is now live on the Vercel Marketplace
Article2026년 5월 26일

Firecrawl is now live on the Vercel Marketplace

Firecrawl이 Vercel Marketplace의 네이티브 통합으로 출시되어, Vercel 프로젝트에서 웹 데이터 수집용 계정·API 키·청구 설정을 자동화할 수 있게 됐습니다.

Eric Ciarla
#agent-memory#agent-routing#context-compression#retrieval-index
On the Shifting Global Compute Landscape
Article2026년 5월 26일

On the Shifting Global Compute Landscape

미국 중심이던 인공지능 연산 생태계가 수출 통제, 중국산 칩의 성장, 개방형 모델과 연산 효율 기술의 확산을 계기로 중국을 포함한 다극적 하드웨어·소프트웨어 구조로 재편되고 있다.

huggingface.co
#nvidia#linear-attention#ai-architecture#agent-memory
Building Convex OS, a Browser-Based React App with Real-Time Sync
Article2026년 5월 23일

Building Convex OS, a Browser-Based React App with Real-Time Sync

Convex OS는 Windows XP 스타일의 브라우저 기반 React 데스크톱 UI를 Convex의 반응형 데이터베이스 상태 모델 위에 올려, 창 위치·프로세스·파일 상태를 여러 탭에서 실시간으로 공유하게 만든 실험이다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
Build a Domain-Specific Embedding Model in Under a Day
Article2026년 5월 20일

Build a Domain-Specific Embedding Model in Under a Day

이 글은 범용 임베딩 모델이 도메인 특유의 미묘한 차이를 놓칠 때, 합성 질의응답 생성·하드 네거티브 마이닝·멀티홉 학습·대조학습·BEIR 평가를 통해 하루 안에 도메인 특화 임베딩 모델로 미세조정하는 절차를 설명한다.

huggingface.co
#nvidia#ai-architecture#multimodal#agent-memory
Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL
Article2026년 5월 19일

Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL

TRL의 델타 가중치 동기화는 연속된 강화학습 단계에서 실제로 바뀐 극소수의 bf16 원소만 희소 세이프텐서로 저장해 허깅페이스 버킷으로 전달함으로써, 대규모 모델의 반복적인 전체 체크포인트 전송 비용과 추론 중단 시간을 크게 줄이는 방식이다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
At aged care provider Regis, AI takes on paperwork so staff can focus on residents
Article2026년 5월 18일

At aged care provider Regis, AI takes on paperwork so staff can focus on residents

호주 노인요양기관 Regis는 RegiCare Assist로 방대한 24시간 보고서와 인수인계 기록을 요약해 임상관리자가 서류보다 입소자 돌봄에 더 많은 시간을 쓰도록 하고 있습니다.

news.microsoft.com
#privacy-design#agent-memory#context-compression#retrieval-index
How to Give AI Agents a Bash Terminal Without Docker or VMs
Article2026년 5월 14일

How to Give AI Agents a Bash Terminal Without Docker or VMs

Convex Sandbox는 Docker나 VM 없이 Convex Node action, just bash, Convex storage를 조합해 AI 에이전트에게 상태가 유지되는 bash 터미널과 가상 파일 시스템을 제공하는 경량 샌드박스입니다.

stack.convex.dev
#agent-memory#agent-routing#capex-cycle#context-compression
Unlocking asynchronicity in continuous batching
Article2026년 5월 14일

Unlocking asynchronicity in continuous batching

연속 배치의 동기식 유휴 시간을 없애기 위해 중앙처리장치의 배치 준비와 그래픽처리장치의 계산을 분리하고, 비기본 CUDA 스트림과 이벤트로 병렬 실행과 작업 순서를 함께 보장하는 방법을 설명한다.

huggingface.co
#agent-routing#llm#semiconductors#applications
이전1…34567…115 / 11다음