Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#llm
Tag1292건YouTube 48Article 1244

#llm

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

Alias / 동의어

large-language-models

연관 태그

#semiconductors공동문서 1040 · 연관도 83%#applications공동문서 974 · 연관도 81%#agent-routing공동문서 512 · 연관도 61%#agent-memory공동문서 468 · 연관도 55%#ai-architecture공동문서 403 · 연관도 51%#privacy-design공동문서 384 · 연관도 50%#context-compression공동문서 342 · 연관도 47%#agent-deployment공동문서 312 · 연관도 47%#service-design공동문서 286 · 연관도 42%#retrieval-index공동문서 213 · 연관도 37%
Previewing GPT-5.6 Sol: a next-generation model
Article2026년 6월 26일

Previewing GPT-5.6 Sol: a next-generation model

OpenAI는 GPT‑5.6 시리즈의 제한적 프리뷰를 시작하며, 대표 모델 Sol과 보급형 모델 Terra·Luna의 성능, 안전장치, 제한 공개 방식, 코딩·생물학·사이버보안 평가 결과를 소개했다.

openai.com
#openai#privacy-design#agent-routing#ai-safety
Production-grade AI agents for financial compliance: Lessons from Stripe
Article2026년 6월 26일

Production-grade AI agents for financial compliance: Lessons from Stripe

Stripe는 대규모 결제·컴플라이언스 검토에서 사람의 최종 판단을 유지하면서, 작업 분해·ReAct 에이전트·전용 에이전트 서비스·LLM 프록시를 결합해 검토 시간을 줄이고 감사 가능성을 확보한 사례를 제시한다.

aws.amazon.com
#privacy-design#service-design#ai-architecture#agent-routing
Run a vLLM Server on HF Jobs in One Command
Article2026년 6월 26일

Run a vLLM Server on HF Jobs in One Command

Hugging Face Jobs에서 vLLM OpenAI 호환 서버를 한 줄 명령으로 띄우고, 토큰 인증으로 호출하며, 필요에 따라 대형 모델·Gradio UI·SSH 디버깅·코딩 에이전트 백엔드까지 확장하는 방법을 설명한다.

huggingface.co
#service-design#ai-architecture#ai-infrastructure#capex-cycle
26: Summer Vibes
Article2026년 6월 26일

26: Summer Vibes

Stratechery의 2026년 26주차 안내문은 AI와 바이브 코딩, 애플의 유럽 내 Siri AI 미출시, 메모리 칩과 중국, 다양한 팟캐스트·기사 콘텐츠를 여름 분위기의 주간 번들로 소개한다.

stratechery.com
#anthropic#agent-memory#context-compression#inflation-risk
Why everyone from OpenAI to SpaceX is building their own chips (and turning up the heat on Nvidia)
Article2026년 6월 26일

Why everyone from OpenAI to SpaceX is building their own chips (and turning up the heat on Nvidia)

OpenAI, Google, Apple, SpaceX 등이 Nvidia 의존을 줄이기 위해 맞춤형 AI 칩을 추진하면서 AI 반도체 시장의 단일 공급자 구조가 흔들리고 있다는 TechCrunch Equity 팟캐스트 소개 글이다.

techcrunch.com
#anthropic#broadcom#nvidia#openai
June 2026: LangChain Newsletter — Fleet On-Call Copilot, Deep Agents Rubrics, and More
Article2026년 6월 25일

June 2026: LangChain Newsletter — Fleet On-Call Copilot, Deep Agents Rubrics, and More

2026년 6월 LangChain 뉴스레터는 온콜 대응, 격리된 컴퓨터 사용, 음성 추적, 자기평가형 에이전트, 배포 교육을 중심으로 에이전트의 개발부터 운영과 개선까지 이어지는 도구와 사례를 소개한다.

langchain.com
#agent-deployment#agent-routing#workflow-automation#llm
Why the Best AI Agents Are Simple: Sierra’s Zack Reneau-Wedeen on the Max Agency Podcast
Article2026년 6월 25일

Why the Best AI Agents Are Simple: Sierra’s Zack Reneau-Wedeen on the Max Agency Podcast

시에라의 제품 책임자 잭 르노 위딘은 최고의 AI 에이전트가 복잡한 다중 에이전트 구조보다 전체 맥락을 가진 단순한 단일 에이전트와 목적에 맞는 모델 조합에서 나온다고 설명한다.

langchain.com
#anthropic#ai-architecture#agent-memory#agent-routing
Adobe acquires image and video enhancement tool maker Topaz Labs
Article2026년 6월 25일

Adobe acquires image and video enhancement tool maker Topaz Labs

Adobe가 AI 기반 이미지·영상 향상 도구 업체 Topaz Labs를 인수해 Firefly와 Creative Cloud 제품군에 Topaz 모델을 통합한다.

Ivan Mehta
#capex-cycle#llm#semiconductors#applications
The Agent Development Lifecycle: Build, Test, Deploy & Monitor AI Agents
Article2026년 6월 25일

The Agent Development Lifecycle: Build, Test, Deploy & Monitor AI Agents

이 글은 AI 에이전트를 일회성 데모가 아니라 반복적으로 구축·검증·배포·관찰하며 개선하는 ‘Agent Development Lifecycle’을 Build, Test, Deploy, Monitor 네 단계로 설명한다.

langchain.com
#agent-deployment#agent-routing#context-compression#prompt-library
Agentic Engineering: How Swarms of AI Agents Are Redefining Software Engineering
Article2026년 6월 25일

Agentic Engineering: How Swarms of AI Agents Are Redefining Software Engineering

이 글은 AI 코딩 도구를 넘어, 역할·기억·관측성을 공유하는 다중 에이전트가 소프트웨어 전달 전 과정을 조율하는 ‘에이전틱 엔지니어링’의 구조와 파일럿 결과를 설명한다.

langchain.com
#agent-swarms#ai-architecture#agent-memory#agent-routing
Why Amazon Dropped Its OpenAI Movie, Data Center Workers Fight Back, and Meta Leaks Employee Data
Article2026년 6월 25일

Why Amazon Dropped Its OpenAI Movie, Data Center Workers Fight Back, and Meta Leaks Employee Data

WIRED ‘Uncanny Valley’는 Amazon MGM의 OpenAI 영화 하차, Google DeepMind와 A24의 AI 협업, 데이터센터 노동자 반발, Meta 직원 추적 프로그램 중단, Anthropic과 정부 관계 변화까지 AI 산업이 문화·노동·감시·정치 영역과 얽히는 방식을 다룬다.

Brian Barrett
#anthropic#openai#meta-ai#privacy-design
Amazon ups India bet with fresh $13B AI infrastructure investment
Article2026년 6월 25일

Amazon ups India bet with fresh $13B AI infrastructure investment

아마존은 2030년까지 인도에서 AI·클라우드 기반을 확대하기 위해 130억 달러를 추가 투자하고, 동시에 현지 물류·퀵커머스 확장에도 속도를 내고 있다.

Jagmeet Singh
#service-design#llm#semiconductors#applications
An Interview with Figma CEO Dylan Field About Design and AI
Article2026년 6월 25일

An Interview with Figma CEO Dylan Field About Design and AI

이 인터뷰는 딜런 필드의 성장 배경과 피그마 창업 과정, 브라우저 기반 협업 디자인 도구로서 피그마가 성립한 기술적 출발점, 그리고 시장이 우려하는 AI를 필드가 기회로 보는 관점을 다룬다.

stratechery.com
#anthropic#capex-cycle#change-management#organizational-redesign
Anthropic's Claude is winning over paid consumers, a market owned by ChatGPT
Article2026년 6월 25일

Anthropic's Claude is winning over paid consumers, a market owned by ChatGPT

앤스로픽의 클로드는 여전히 챗지피티보다 훨씬 작지만, 결제 데이터와 교육 수요 지표에서 유료 소비자층이 빠르게 확대되고 있는 것으로 나타났다.

Julie Bort
#anthropic#capex-cycle#llm#semiconductors
British Police Built a Sprawling Crime-Prediction Machine. Some Results Couldn’t Be Trusted
Article2026년 6월 25일

British Police Built a Sprawling Crime-Prediction Machine. Some Results Couldn’t Be Trusted

WIRED 조사에 따르면 영국 브리스틀 지역 경찰과 시의회는 민감한 공공 데이터를 결합해 대규모 예측 치안·위험 점수 시스템을 만들었지만, 일부 모델은 신뢰성·투명성 문제로 폐기되거나 비판을 받았다.

wired.com
#privacy-design#service-design#travel-hospitality#llm
Building agentic AI applications with a modern data mesh strategy on AWS
Article2026년 6월 25일

Building agentic AI applications with a modern data mesh strategy on AWS

이 글은 RAG의 단일 검색 지점 통제를 넘어, 에이전트형 AI가 스키마 탐색·SQL 생성·쿼리 실행·응답 합성까지 수행하는 전 과정에 세밀한 권한 통제를 적용하기 위한 AWS 기반 서버리스 데이터 메시 아키텍처를 설명한다.

aws.amazon.com
#service-design#ai-architecture#agent-memory#context-compression
Databricks’ former AI chief thinks he can cut AI’s power bill by 1,000x
Article2026년 6월 25일

Databricks’ former AI chief thinks he can cut AI’s power bill by 1,000x

Databricks 전 AI 책임자 Naveen Rao가 이끄는 Unconventional AI는 오실레이터 기반 컴퓨팅으로 AI 추론 전력 사용을 최대 1,000분의 1로 줄이겠다는 목표를 제시했다.

Russell Brandom
#ai-architecture#capex-cycle#llm#applications
General Intuition's $2.3B bet that video games can train AI agents for the real world
Article2026년 6월 25일

General Intuition's $2.3B bet that video games can train AI agents for the real world

General Intuition은 게임 플레이 데이터와 입력 행동 기록을 기반으로 현실 세계에서 작동할 수 있는 범용 AI 에이전트를 훈련시키겠다는 목표로 23억 달러 가치 평가를 받았다.

Rebecca Bellan
#anthropic#capex-cycle#inflation-risk#llm
How LangSmith and LangChain OSS Help You Meet EU AI Act Requirements
Article2026년 6월 25일

How LangSmith and LangChain OSS Help You Meet EU AI Act Requirements

이 글은 EU AI Act의 고위험 AI 시스템 요구사항을 운영 인프라 관점에서 해석하고, LangSmith와 LangChain OSS가 추적, 평가, 인간 감독, 데이터 거주성 측면에서 이를 어떻게 지원하는지 설명한다.

langchain.com
#ai-architecture#agent-deployment#agent-routing#prompt-library
How Klarna's AI assistant redefined customer support at scale for 85 million active users
Article2026년 6월 25일

How Klarna's AI assistant redefined customer support at scale for 85 million active users

Klarna는 LangGraph와 LangSmith 기반 AI Assistant로 결제·환불·에스컬레이션 업무를 대규모로 처리하며 고객지원 속도와 자동화 수준을 크게 높였다.

langchain.com
#service-design#ai-architecture#context-compression#prompt-library
Optimize model training on Amazon SageMaker AI with NVIDIA Blackwell
Article2026년 6월 25일

Optimize model training on Amazon SageMaker AI with NVIDIA Blackwell

이 글은 Amazon SageMaker AI의 P6 B200 인스턴스와 NVIDIA Blackwell GPU를 활용해 대규모 모델 학습에서 배치 크기, 시퀀스 길이, 샤딩, 정밀도 형식, 활성화 체크포인팅을 어떻게 조정해야 하는지 설명한다.

aws.amazon.com
#nvidia#service-design#ai-architecture#agent-memory
Optimizing cloud economics with linear elastic caching
Article2026년 6월 25일

Optimizing cloud economics with linear elastic caching

선형 탄력 캐싱은 캐시 메모리 비용과 캐시 미스 비용을 함께 최적화하기 위해 페이지 보존 시간을 스키 대여 문제로 모델링하고, 가벼운 학습 기반 TTL 예측으로 고정 크기 캐시보다 낮은 총소유비용을 달성하려는 접근이다.

research.google
#resort-experience#privacy-design#service-design#agent-memory
Patronus AI lands $50M to build ‘digital worlds’ that stress-test AI agents
Article2026년 6월 25일

Patronus AI lands $50M to build ‘digital worlds’ that stress-test AI agents

Patronus AI는 AI 에이전트가 실제 업무를 맡기 전에 복잡한 상황에서 제대로 작동하는지 검증하는 ‘디지털 세계’ 시뮬레이션을 만들며 5,000만 달러 규모의 시리즈 B 투자를 유치했다.

Marina Temkin
#capex-cycle#llm#semiconductors#applications
Repositioning retail for the AI era
Article2026년 6월 25일

Repositioning retail for the AI era

Macy’s는 소매업에서 인공지능을 겉으로 드러나는 기능보다 검색, 재고, 운영, 개발 의사결정에 내재화하는 ‘AI first’ 전략으로 재편하고 있다.

technologyreview.com
#agent-deployment#agent-routing#llm#semiconductors
이전1…910111213…5411 / 54다음