Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#agent-memory
Tag543건YouTube 19Article 524

#agent-memory

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#retrieval-index공동문서 248 · 연관도 66%#context-compression공동문서 286 · 연관도 60%#semiconductors공동문서 457 · 연관도 56%#applications공동문서 432 · 연관도 56%#llm공동문서 468 · 연관도 55%#agent-routing공동문서 201 · 연관도 37%#ai-architecture공동문서 152 · 연관도 30%#agent-deployment공동문서 98 · 연관도 23%#vision-language-models공동문서 64 · 연관도 20%#multimodal공동문서 66 · 연관도 20%
Firecrawl is now live on the Vercel Marketplace
Article2026년 5월 26일

Firecrawl is now live on the Vercel Marketplace

Firecrawl이 Vercel Marketplace의 네이티브 통합으로 출시되어, Vercel 프로젝트에서 웹 데이터 수집용 계정·API 키·청구 설정을 자동화할 수 있게 됐습니다.

Eric Ciarla
#agent-memory#agent-routing#context-compression#retrieval-index
On the Shifting Global Compute Landscape
Article2026년 5월 26일

On the Shifting Global Compute Landscape

미국 중심이던 인공지능 연산 생태계가 수출 통제, 중국산 칩의 성장, 개방형 모델과 연산 효율 기술의 확산을 계기로 중국을 포함한 다극적 하드웨어·소프트웨어 구조로 재편되고 있다.

huggingface.co
#nvidia#linear-attention#ai-architecture#agent-memory
Building Convex OS, a Browser-Based React App with Real-Time Sync
Article2026년 5월 23일

Building Convex OS, a Browser-Based React App with Real-Time Sync

Convex OS는 Windows XP 스타일의 브라우저 기반 React 데스크톱 UI를 Convex의 반응형 데이터베이스 상태 모델 위에 올려, 창 위치·프로세스·파일 상태를 여러 탭에서 실시간으로 공유하게 만든 실험이다.

stack.convex.dev
#agent-routing#llm#semiconductors#applications
Issue 354
Article2026년 5월 22일

Issue 354

원문은 하버드의 A 학점 제한을 비판하며 교육의 목적을 ‘판정’보다 ‘성공 지원’에 두어야 한다고 주장하고, 이어 Hermes Agent의 OpenClaw 추격과 실시간 상호작용형 멀티모달 모델 TML Interaction Small을 소개한다.

@DeepLearningAI
#openclaw#anthropic#token-efficiency#ai-architecture
LeRobot Humanoid: An Open, Low-Cost, 3D-Printed Humanoid for Robot Learning
Article2026년 5월 21일

LeRobot Humanoid: An Open, Low-Cost, 3D-Printed Humanoid for Robot Learning

LeRobot Humanoid는 약 2,500달러 부품비의 3D 프린트 기반 개방형 휴머노이드 플랫폼으로, 하드웨어 제작부터 시뮬레이션, 식별, 학습, 실제 제어까지 로봇 학습 전 과정을 재현 가능하게 만들려는 프로젝트다.

Hugging Face
#lerobot#mjlab#hugging-face#lerobot-humanoid
All Systems Nominal: shaping the future of hardware development
Article2026년 5월 21일

All Systems Nominal: shaping the future of hardware development

전 해군 장교 캐머런 맥코드가 잠수함·의회·방산 스타트업 현장에서 겪은 복잡한 시스템 운용 경험은, 실제 배치되는 하드웨어를 더 빠르고 안전하게 시험하려는 Nominal의 문제의식으로 이어졌다.

sequoiacap.com
#service-design#llm#applications#robotics
New Approach to Scaling Laws Could Change How AI Models Are Trained
Article2026년 5월 21일

New Approach to Scaling Laws Could Change How AI Models Are Trained

스탠퍼드 연구진은 교육측정학의 문항응답 원리를 스케일링 법칙에 적용해, 대형 언어모델 성능 예측에 필요한 계산량을 크게 줄이는 아이템 응답 스케일링 법칙을 제안했다.

hai.stanford.edu
#privacy-design#ai-architecture#llm#semiconductors
What senior leaders want to know about AI
Article2026년 5월 20일

What senior leaders want to know about AI

MIT Sloan의 피터 허스트는 고위 리더들이 AI를 단순한 기술 문제가 아니라 조직, 사람, IT와의 관계, 위험 관리, 실행 역량의 문제로 이해하려 한다고 설명한다.

mitsloan.mit.edu
#agent-routing#change-management#organizational-redesign#semiconductors
Build a Domain-Specific Embedding Model in Under a Day
Article2026년 5월 20일

Build a Domain-Specific Embedding Model in Under a Day

이 글은 범용 임베딩 모델이 도메인 특유의 미묘한 차이를 놓칠 때, 합성 질의응답 생성·하드 네거티브 마이닝·멀티홉 학습·대조학습·BEIR 평가를 통해 하루 안에 도메인 특화 임베딩 모델로 미세조정하는 절차를 설명한다.

huggingface.co
#nvidia#ai-architecture#multimodal#agent-memory
Everything new in our Google AI subscriptions, fresh from I/O 2026
Article2026년 5월 19일

Everything new in our Google AI subscriptions, fresh from I/O 2026

Google은 I/O 2026에서 새로운 월 100달러 AI Ultra 요금제, 기존 최상위 Ultra 가격 인하, Gemini Omni·Gemini 3.5 Flash·Gemini Spark·Project Genie 접근 확대, Gmail·Gemini 앱 생산성 기능과 compute 기반 사용량 체계를 포함한 Google AI 구독 개편을 발표했다.

blog.google
#agent-routing#context-compression#prompt-library#search-advertising
OlmoEarth v1.1: A more efficient family of Earth observation models
Article2026년 5월 19일

OlmoEarth v1.1: A more efficient family of Earth observation models

OlmoEarth v1.1은 위성영상 토큰의 해상도별 분리를 통합하고 사전학습 방식을 조정해 기존 성능을 대체로 유지하면서 연산 비용을 최대 3분의 1로 줄인 지구관측 모델 제품군이다.

huggingface.co
#ai-architecture#llm#semiconductors#applications
Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL
Article2026년 5월 19일

Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL

TRL의 델타 가중치 동기화는 연속된 강화학습 단계에서 실제로 바뀐 극소수의 bf16 원소만 희소 세이프텐서로 저장해 허깅페이스 버킷으로 전달함으로써, 대규모 모델의 반복적인 전체 체크포인트 전송 비용과 추론 중단 시간을 크게 줄이는 방식이다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#context-compression
At aged care provider Regis, AI takes on paperwork so staff can focus on residents
Article2026년 5월 18일

At aged care provider Regis, AI takes on paperwork so staff can focus on residents

호주 노인요양기관 Regis는 RegiCare Assist로 방대한 24시간 보고서와 인수인계 기록을 요약해 임상관리자가 서류보다 입소자 돌봄에 더 많은 시간을 쓰도록 하고 있습니다.

news.microsoft.com
#privacy-design#agent-memory#context-compression#retrieval-index
Learn the Hugging Face Kernel Hub in 5 Minutes
Article2026년 5월 18일

Learn the Hugging Face Kernel Hub in 5 Minutes

허깅페이스 커널 허브는 파이썬·파이토치·쿠다 환경에 맞는 사전 최적화 연산 커널을 허브에서 내려받아, 복잡한 로컬 빌드 없이 모델의 특정 연산에 적용하도록 돕는 체계다.

huggingface.co
#ai-architecture#agent-routing#llm#applications
Resilient AI End-to-End Tests with Stagehand and Convex
Article2026년 5월 16일

Resilient AI End-to-End Tests with Stagehand and Convex

이 글은 Stagehand의 자연어 브라우저 제어와 Convex의 실행별 임시 백엔드를 결합해, 선택자 중심 E2E 테스트의 취약성을 줄이고 실제 사용자 의도에 가까운 테스트를 구성한 경험을 정리한다.

stack.convex.dev
#agent-deployment#agent-routing#llm#semiconductors
How to Give AI Agents a Bash Terminal Without Docker or VMs
Article2026년 5월 14일

How to Give AI Agents a Bash Terminal Without Docker or VMs

Convex Sandbox는 Docker나 VM 없이 Convex Node action, just bash, Convex storage를 조합해 AI 에이전트에게 상태가 유지되는 bash 터미널과 가상 파일 시스템을 제공하는 경량 샌드박스입니다.

stack.convex.dev
#agent-memory#agent-routing#capex-cycle#context-compression
Unlocking asynchronicity in continuous batching
Article2026년 5월 14일

Unlocking asynchronicity in continuous batching

연속 배치의 동기식 유휴 시간을 없애기 위해 중앙처리장치의 배치 준비와 그래픽처리장치의 계산을 분리하고, 비기본 CUDA 스트림과 이벤트로 병렬 실행과 작업 순서를 함께 보장하는 방법을 설명한다.

huggingface.co
#agent-routing#llm#semiconductors#applications
Introducing Langsmith Engine
Article2026년 5월 13일

Introducing Langsmith Engine

LangSmith Engine은 운영 트레이스에서 반복되는 에이전트 실패를 자동으로 묶고, 원인을 진단하며, 수정 PR과 평가 커버리지를 제안하는 LangSmith의 공개 베타 기능입니다.

langchain.com
#agent-routing#context-compression#prompt-library#semiconductors
How ‘learnrights’ would compensate creators for AI model training
Article2026년 5월 12일

How ‘learnrights’ would compensate creators for AI model training

MIT 슬론의 토머스 말론과 공동 저자들은 생성형 AI 학습에 쓰이는 저작물에 대해 창작자가 라이선스 권리를 갖고 보상받도록 하는 ‘learnright’ 제도를 제안한다.

mitsloan.mit.edu
#anthropic#privacy-design#llm#semiconductors
AutoScout24 scales engineering with AI-powered workflows
Article2026년 5월 12일

AutoScout24 scales engineering with AI-powered workflows

AutoScout24는 조직 전반의 ChatGPT 도입과 엔지니어링 워크플로에 통합된 Codex 활용을 병행해, 복잡해지는 제품·시스템 환경 속에서 개발 속도와 품질을 함께 높이려는 AI 전환을 추진했다.

openai.com
#autoscout24#chatgpt#openai#autotrader-ca
Staying ahead in the age of AI
Article2026년 5월 11일

Staying ahead in the age of AI

OpenAI의 「Staying ahead in the age of AI」는 기업이 AI 발전 속도에 뒤처지지 않기 위해 정렬, 활성화, 확산, 가속, 거버넌스라는 실행 원칙을 조직 운영에 통합해야 한다고 설명하는 리더십 가이드입니다.

openai.com
#agent-routing#ai-safety#llm#semiconductors
Why Convex doesn't let candidates use AI in coding interviews
Article2026년 5월 11일

Why Convex doesn't let candidates use AI in coding interviews

Convex는 AI를 부정해서가 아니라 코딩 인터뷰가 지원자의 사고·소통·판단을 읽는 제한된 장치이기 때문에 AI를 배제하고, 실제 업무에서는 산출물 품질에 책임지는 한 도구 사용을 강제도 금지도 하지 않는다.

stack.convex.dev
#ai-architecture#agent-routing#llm#semiconductors
Testing ads in ChatGPT
Article2026년 5월 7일

Testing ads in ChatGPT

OpenAI는 ChatGPT의 무료·저가 접근성을 유지하기 위해 광고 파일럿을 확대하되, 답변 독립성·대화 프라이버시·사용자 통제를 핵심 원칙으로 유지한다고 밝혔다.

openai.com
#openai#privacy-design#agent-memory#context-compression
vLLM V0 to V1: Correctness Before Corrections in RL
Article2026년 5월 6일

vLLM V0 to V1: Correctness Before Corrections in RL

온라인 강화학습의 vLLM V0→V1 전환에서는 목적함수 보정을 서두르기보다 로그확률 의미, 런타임 기본값, 비행 중 가중치 갱신, fp32 출력 헤드를 먼저 일치시켜 추론 백엔드의 정확성을 복원해야 했다.

huggingface.co
#llm#semiconductors#applications#agent-deployment
이전1…89101112…2310 / 23다음