Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#retrieval-index
Tag248건YouTube 17Article 231

#retrieval-index

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-memory공동문서 248 · 연관도 66%#context-compression공동문서 160 · 연관도 50%#applications공동문서 197 · 연관도 38%#llm공동문서 213 · 연관도 37%#semiconductors공동문서 203 · 연관도 37%#agent-routing공동문서 103 · 연관도 28%#ai-architecture공동문서 82 · 연관도 24%#service-design공동문서 41 · 연관도 14%#gpu공동문서 17 · 연관도 12%#multimodal공동문서 25 · 연관도 11%
Testing ads in ChatGPT
Article2026년 5월 7일

Testing ads in ChatGPT

OpenAI는 ChatGPT의 무료·저가 접근성을 유지하기 위해 광고 파일럿을 확대하되, 답변 독립성·대화 프라이버시·사용자 통제를 핵심 원칙으로 유지한다고 밝혔다.

openai.com
#openai#privacy-design#agent-memory#context-compression
vLLM V0 to V1: Correctness Before Corrections in RL
Article2026년 5월 6일

vLLM V0 to V1: Correctness Before Corrections in RL

온라인 강화학습의 vLLM V0→V1 전환에서는 목적함수 보정을 서두르기보다 로그확률 의미, 런타임 기본값, 비행 중 가중치 갱신, fp32 출력 헤드를 먼저 일치시켜 추론 백엔드의 정확성을 복원해야 했다.

huggingface.co
#llm#semiconductors#applications#agent-deployment
Granite Embedding Multilingual R2: Open Apache 2.0 Multilingual Embeddings with 32K Context — Best Sub-100M Retrieval Quality
Article2026년 5월 4일

Granite Embedding Multilingual R2: Open Apache 2.0 Multilingual Embeddings with 32K Context — Best Sub-100M Retrieval Quality

IBM의 Granite Embedding Multilingual R2는 200개 이상의 언어와 32K 토큰 문맥을 지원하며, 97M 소형 모델과 311M 고성능 모델로 다국어·장문·코드 검색의 성능과 배포 효율을 함께 높인 Apache 2.0 임베딩 모델군이다.

huggingface.co
#ai-architecture#multimodal#agent-memory#retrieval-index
Why Stanford Is Restructuring For AI’s Next Era
Article2026년 5월 4일

Why Stanford Is Restructuring For AI’s Next Era

스탠퍼드는 AI 변화 속도에 맞춰 HAI와 스탠퍼드 데이터 사이언스를 통합하고, 대규모 협업 연구와 개방성을 중심으로 대학의 AI 역할을 재구성하려 한다.

hai.stanford.edu
#service-design#agent-routing#llm#semiconductors
Introducing /parse: Turn any document into LLM-ready data
Article2026년 4월 28일

Introducing /parse: Turn any document into LLM-ready data

Firecrawl은 로컬 문서를 업로드해 웹 페이지처럼 정리된 마크다운, 요약, 구조화 JSON을 받을 수 있는 새 API /parse를 출시했다.

Eric Ciarla
#agent-memory#agent-routing#capex-cycle#retrieval-index
It's all about the angle: Your photos, re-composed
Article2026년 4월 22일

It's all about the angle: Your photos, re-composed

Google Photos의 Auto frame은 촬영 후에도 사진을 3D 장면처럼 해석해 시점과 구도를 자동으로 재구성하는 이미지 편집 기능이다.

research.google
#llm#semiconductors#applications#agent-memory
Scaling Codex to enterprises worldwide
Article2026년 4월 21일

Scaling Codex to enterprises worldwide

Codex는 주간 개발자 사용자가 2주 만에 300만 명 이상에서 400만 명 이상으로 늘어난 가운데, 기업들이 소프트웨어 개발과 지식 업무 전반의 실제 워크플로에 빠르게 도입하는 단계로 확장되고 있다.

openai.com
#agent-memory#agent-routing#context-compression#retrieval-index
Fragments: April 14
Article2026년 4월 14일

Fragments: April 14

마틴 파울러는 AI 시대에도 좋은 소프트웨어를 만드는 핵심은 더 많은 코드를 빠르게 생산하는 능력이 아니라, 인간의 ‘게으름’이 낳는 단순한 추상화와 의심·검증·절제라고 정리한다.

martinfowler.com
#agent-memory#retrieval-index#llm#semiconductors
Fragments: April 9
Article2026년 4월 9일

Fragments: April 9

마틴 파울러는 두 편의 팟캐스트, 공급망 공격 사례, 문서화 프레임워크, AI 에이전트 기반 개발 경험, 돌봄 중심의 성장 관점을 묶어 소개한다.

martinfowler.com
#ai-architecture#agent-memory#change-management#organizational-redesign
Five Big Improvements to Gradio MCP Servers
Article2026년 4월 6일

Five Big Improvements to Gradio MCP Servers

Gradio 5.38.0은 로컬 파일 전달, 실시간 진행 알림, OpenAPI 변환, 인증 헤더 처리, 도구 설명 맞춤화로 MCP 서버의 개발·연결 경험을 개선했다.

huggingface.co
#agent-routing#llm#semiconductors#applications
Fragments: April 2
Article2026년 4월 2일

Fragments: April 2

마틴 파울러는 LLM 시대의 핵심 문제가 코드 생산 자체보다 시스템 이해, 의도 보존, 검증, 그리고 인간과 AI가 함께 사용할 언어의 설계로 이동하고 있다고 정리한다.

martinfowler.com
#agent-memory#energy-security#inflation-risk#retrieval-index
Firecrawl + n8n: Bring Real-Time Web Data Into Your AI Workflows
Article2026년 3월 26일

Firecrawl + n8n: Bring Real-Time Web Data Into Your AI Workflows

Firecrawl과 n8n의 네이티브 통합은 웹페이지를 LLM이 바로 쓸 수 있는 정제 데이터로 바꾸고, 이를 AI 워크플로·RAG·리드 강화·시장 조사 자동화에 연결하도록 해준다.

Nicolas Camara
#agent-memory#agent-routing#context-compression#retrieval-index
Introducing /interact — Scrape and interact with a web page
Article2026년 3월 25일

Introducing /interact — Scrape and interact with a web page

Firecrawl의 /interact는 스크레이프한 웹페이지 세션을 그대로 살아 있는 브라우저로 전환해 클릭, 입력, 이동, 동적 데이터 추출을 수행하게 하는 새 엔드포인트다.

Eric Ciarla
#agent-memory#agent-routing#context-compression#prompt-library
TurboQuant: Redefining AI efficiency with extreme compression
Article2026년 3월 24일

TurboQuant: Redefining AI efficiency with extreme compression

TurboQuant는 QJL과 PolarQuant를 결합해 벡터 양자화의 메모리 오버헤드를 줄이고, KV 캐시 압축과 고차원 벡터 검색에서 정확도 손실 없이 큰 압축·속도 향상을 보인 Google Research의 알고리즘입니다.

research.google
#privacy-design#agent-memory#context-compression#retrieval-index
Holotron-12B - High Throughput Computer Use Agent
Article2026년 3월 17일

Holotron-12B - High Throughput Computer Use Agent

홀로트론 12B는 하이브리드 상태공간모델과 어텐션 구조를 기반으로 긴 문맥과 다중 이미지를 효율적으로 처리하면서 높은 컴퓨터 사용 성능과 추론 처리량을 달성한 120억 매개변수 멀티모달 에이전트 모델이다.

huggingface.co
#nvidia#ai-architecture#multimodal#agent-memory
Introducing Groundsource: Turning news reports into data with Gemini
Article2026년 3월 12일

Introducing Groundsource: Turning news reports into data with Gemini

Google Research는 Gemini로 전 세계 뉴스 보도를 분석해 도시 돌발홍수 260만 건의 구조화 데이터셋을 만든 Groundsource를 공개하고, 이를 예측·연구 기반으로 활용한다고 밝혔다.

research.google
#llm#semiconductors#applications#agent-memory
Exploring the feasibility of conversational diagnostic AI in a real-world clinical study
Article2026년 3월 11일

Exploring the feasibility of conversational diagnostic AI in a real-world clinical study

구글 리서치·구글 딥마인드와 Beth Israel Deaconess Medical Center의 전향적 단일기관 연구는 대화형 의료 AI AMIE가 실제 일차진료 방문 전 병력 청취를 감독하에 수행하는 것이 초기 단계에서 실행 가능하고 안전하게 수용될 수 있음을 보고했다.

research.google
#multimodal#agent-routing#semiconductors#vision-language-models
Introducing Storage Buckets on the Hugging Face Hub
Article2026년 3월 11일

Introducing Storage Buckets on the Hugging Face Hub

Hugging Face Storage Buckets는 자주 변경되는 머신러닝 중간 산출물을 빠르게 저장·동기화·관리하고, Xet 기반 중복 제거와 사전 워밍을 통해 전송 효율과 데이터 접근성을 높이는 Hub 네이티브 객체 스토리지입니다.

huggingface.co
#service-design#agent-memory#agent-routing#retrieval-index
Partnering with Scanner: Every Log Tells a Story—If You Can Find It Fast Enough
Article2026년 3월 10일

Partnering with Scanner: Every Log Tells a Story—If You Can Find It Fast Enough

Sequoia는 Scanner가 저비용 객체 스토리지에 묻힌 방대한 보안 로그를 초 단위로 검색 가능하게 만들어, AI 기반 보안 운영의 핵심 인프라가 되고 있다고 설명한다.

sequoiacap.com
#agent-routing#semiconductors#applications#compute
Where wild things roam: Identifying wildlife with SpeciesNet
Article2026년 3월 6일

Where wild things roam: Identifying wildlife with SpeciesNet

Google Research가 공개한 오픈소스 AI 모델 SpeciesNet은 카메라 트랩 이미지 속 야생동물을 대규모로 자동 식별해 연구와 보전 활동의 속도와 범위를 넓히고 있다.

research.google
#agent-routing#llm#semiconductors#applications
Ideological Resistance to Patents, Followed by Reluctant Pragmatism
Article2026년 3월 5일

Ideological Resistance to Patents, Followed by Reluctant Pragmatism

이 글은 소프트웨어 특허에 대한 이념적 반감이 실제 특허 공격 경험과 스타트업의 법적 비대칭성을 거치며 방어적 특허라는 현실적 선택으로 바뀌는 과정을 설명한다.

martinfowler.com
#ai-architecture#agent-memory#retrieval-index#llm
Tricks from OpenAI gpt-oss YOU 🫵 can use with transformers
Article2026년 3월 5일

Tricks from OpenAI gpt-oss YOU 🫵 can use with transformers

OpenAI의 GPT OSS 모델을 transformers에서 효율적으로 실행·미세조정하기 위해 도입된 다운로드형 커널, MXFP4 양자화, Flash Attention 3, 텐서 병렬화 등 핵심 업그레이드를 설명한 글입니다.

huggingface.co
#openai#ai-architecture#agent-memory#agent-routing
When to and when not to use return validators
Article2026년 2월 27일

When to and when not to use return validators

Convex의 return validator는 유용하지만 모든 쿼리·뮤테이션에 기계적으로 붙일 도구가 아니라, 정확한 런타임 반환 계약이 필요할 때 선택적으로 써야 한다는 지침 변경이다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
Introducing PDF Parser v2: Faster Extraction with Auto Mode
Article2026년 2월 26일

Introducing PDF Parser v2: Faster Extraction with Auto Mode

Firecrawl은 PDF Parser v2를 공개하며 Rust 기반 파서, 기본 Auto 모드, OCR 자동 전환을 통해 웹 PDF를 더 빠르고 안정적으로 구조화 데이터로 추출할 수 있게 했습니다.

Eric Ciarla
#llm#semiconductors#applications#agent-deployment
이전1…45678…116 / 11다음