Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#llm
Tag1292건YouTube 48Article 1244

#llm

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

Alias / 동의어

large-language-models

연관 태그

#semiconductors공동문서 1040 · 연관도 83%#applications공동문서 974 · 연관도 81%#agent-routing공동문서 512 · 연관도 61%#agent-memory공동문서 468 · 연관도 55%#ai-architecture공동문서 403 · 연관도 51%#privacy-design공동문서 384 · 연관도 50%#context-compression공동문서 342 · 연관도 47%#agent-deployment공동문서 312 · 연관도 47%#service-design공동문서 286 · 연관도 42%#retrieval-index공동문서 213 · 연관도 37%
Staying ahead in the age of AI
Article2026년 5월 11일

Staying ahead in the age of AI

OpenAI의 「Staying ahead in the age of AI」는 기업이 AI 발전 속도에 뒤처지지 않기 위해 정렬, 활성화, 확산, 가속, 거버넌스라는 실행 원칙을 조직 운영에 통합해야 한다고 설명하는 리더십 가이드입니다.

openai.com
#agent-routing#ai-safety#llm#semiconductors
Why Convex doesn't let candidates use AI in coding interviews
Article2026년 5월 11일

Why Convex doesn't let candidates use AI in coding interviews

Convex는 AI를 부정해서가 아니라 코딩 인터뷰가 지원자의 사고·소통·판단을 읽는 제한된 장치이기 때문에 AI를 배제하고, 실제 업무에서는 산출물 품질에 책임지는 한 도구 사용을 강제도 금지도 하지 않는다.

stack.convex.dev
#ai-architecture#agent-routing#llm#semiconductors
Introducing Question and Highlights: High-Quality Answers from the Web, 100x Fewer Tokens
Article2026년 5월 8일

Introducing Question and Highlights: High-Quality Answers from the Web, 100x Fewer Tokens

Firecrawl은 /scrape에 question과 highlights 형식을 추가해, 전체 페이지를 긁어 LLM에 넣는 복잡한 과정을 한 번의 호출로 줄이고 URL 기반의 근거 있는 답변이나 원문 그대로의 발췌를 훨씬 적은 토큰으로 반환한다고 발표했다.

Eric Ciarla
#token-efficiency#agent-routing#context-compression#prompt-library
Running Codex safely at OpenAI
Article2026년 5월 8일

Running Codex safely at OpenAI

OpenAI는 Codex를 제한된 실행 환경, 승인 정책, 관리형 네트워크·인증 통제, 에이전트 친화적 텔레메트리로 운영해 개발 생산성과 보안 가시성을 함께 확보한다고 설명한다.

openai.com
#openai#privacy-design#agent-routing#context-compression
Seedance Makes A Splash, Nvidia's AI-Guided Chip Designs, Helping Robots Not Forget
Article2026년 5월 8일

Seedance Makes A Splash, Nvidia's AI-Guided Chip Designs, Helping Robots Not Forget

본문은 AI 대량실업론에 대한 반박으로 시작해 ByteDance의 Seedance 2.0 확산과 Nvidia의 AI 기반 칩 설계 활용 사례를 다룬다.

deeplearning.ai
#nvidia#privacy-design#service-design#ai-architecture
Inside Anthropic’s 2026 Developer Conference
Article2026년 5월 7일

Inside Anthropic’s 2026 Developer Conference

Anthropic의 2026 개발자 콘퍼런스에서 가장 큰 발표는 새 모델이 아니라 SpaceX의 Colossus 슈퍼클러스터 용량을 Claude에 배정하는 컴퓨트 계약과, Claude Managed Agents를 중심으로 한 플랫폼 전환이었다.

Every
#anthropic#claude-code#spacex-colossus#claude-managed-agents
Inside Porsche Cup Brasil’s AI-powered race operations
Article2026년 5월 7일

Inside Porsche Cup Brasil’s AI-powered race operations

포르쉐 컵 브라질은 Microsoft 기반 AI 손상 분석과 실시간 텔레메트리를 활용해 사고 차량 진단, 수리 의사결정, 경기 운영을 더 빠르고 일관된 실시간 시스템으로 바꾸고 있다.

news.microsoft.com
#privacy-design#service-design#ai-architecture#agent-routing
Notes from inside China's AI labs - by Nathan Lambert
Article2026년 5월 7일

Notes from inside China's AI labs - by Nathan Lambert

중국의 주요 AI 연구소들은 미국과 비슷한 기술 재료를 갖췄지만, 학생 중심 인력 구조와 낮은 개인주의적 경쟁, 실용적 실행 문화 덕분에 LLM ‘빠른 추격자’로 강하게 기능하고 있다는 현장 관찰이다.

interconnects.ai
#anthropic#china#hotel-review#privacy-design
Parloa builds service agents customers want to talk to
Article2026년 5월 7일

Parloa builds service agents customers want to talk to

Parloa는 OpenAI 모델을 활용해 기업용 음성 고객서비스 에이전트를 설계, 시뮬레이션, 평가, 운영하는 AMP 플랫폼을 구축하고 있으며, 실제 운영 환경에서의 신뢰성·지연시간·일관성을 핵심 기준으로 삼고 있다.

openai.com
#openai#parloa#speech-recognition#gpt-5-4
Scaling Trusted Access for Cyber with GPT-5.5 and GPT-5.5-Cyber
Article2026년 5월 7일

Scaling Trusted Access for Cyber with GPT-5.5 and GPT-5.5-Cyber

OpenAI는 GPT‑5.5와 제한 preview인 GPT‑5.5‑Cyber를 통해 검증된 사이버 방어자에게 더 유용한 접근을 제공하되, 신원 기반 통제와 오용 방지 장치를 결합해 방어 생태계 전반의 대응 속도를 높이려 한다.

openai.com
#openai#privacy-design#service-design#agent-deployment
Simplex rethinks software development with Codex
Article2026년 5월 7일

Simplex rethinks software development with Codex

Simplex는 ChatGPT Enterprise와 Codex를 전사적으로 도입해 AI 기반 개발 방식을 검증하고, 설계·구현·테스트 전반에서 생산성 향상을 정량적으로 확인했다.

openai.com
#openai#privacy-design#agent-routing#llm
Testing ads in ChatGPT
Article2026년 5월 7일

Testing ads in ChatGPT

OpenAI는 ChatGPT의 무료·저가 접근성을 유지하기 위해 광고 파일럿을 확대하되, 답변 독립성·대화 프라이버시·사용자 통제를 핵심 원칙으로 유지한다고 밝혔다.

openai.com
#openai#privacy-design#agent-memory#context-compression
How frontier firms are pulling ahead
Article2026년 5월 6일

How frontier firms are pulling ahead

OpenAI는 B2B Signals를 통해 AI를 더 깊고 넓게, 더 위임된 워크플로에 쓰는 프런티어 기업들이 일반 기업보다 빠르게 앞서고 있으며 그 격차가 누적되기 시작했다고 설명한다.

openai.com
#openai#privacy-design#agent-routing#ai-safety
Introducing ChatGPT Futures: Class of 2026
Article2026년 5월 6일

Introducing ChatGPT Futures: Class of 2026

OpenAI는 AI를 책임감 있고 창의적으로 활용해 실제 변화를 만들고 있는 학생·젊은 빌더 26명을 ‘ChatGPT Futures: Class of 2026’ 첫 기수로 소개했다.

openai.com
#openai#privacy-design#context-compression#prompt-library
Singular Bank helps bankers move fast with ChatGPT and Codex
Article2026년 5월 6일

Singular Bank helps bankers move fast with ChatGPT and Codex

마드리드의 프라이빗 뱅크 Singular Bank는 ChatGPT와 Codex 기반 내부 도우미 Singularity로 포트폴리오 분석, 회의 준비, 후속 커뮤니케이션을 자동화해 은행원들이 하루 60~90분을 절약하도록 했다.

openai.com
#openai#privacy-design#agent-routing#llm
Uber uses OpenAI to help people earn smarter and book faster
Article2026년 5월 6일

Uber uses OpenAI to help people earn smarter and book faster

우버는 OpenAI 모델과 Realtime API를 활용해 운전자에게 실시간 수익 기회를 더 쉽게 안내하고, 승객에게는 자연어 음성 기반 예약 경험을 제공하려 하고 있다.

openai.com
#openai#privacy-design#ai-architecture#agent-routing
vLLM V0 to V1: Correctness Before Corrections in RL
Article2026년 5월 6일

vLLM V0 to V1: Correctness Before Corrections in RL

온라인 강화학습의 vLLM V0→V1 전환에서는 목적함수 보정을 서두르기보다 로그확률 의미, 런타임 기본값, 비행 중 가중치 갱신, fp32 출력 헤드를 먼저 일치시켜 추론 백엔드의 정확성을 복원해야 했다.

huggingface.co
#llm#semiconductors#applications#agent-deployment
Advancing youth safety and wellbeing in EMEA
Article2026년 5월 5일

Advancing youth safety and wellbeing in EMEA

OpenAI는 유럽 청소년 안전 청사진과 EMEA 청소년·웰빙 그랜트 첫 수혜자를 발표하며, 청소년이 AI를 안전하고 발달에 맞게 활용하도록 정책 권고와 현장 지원을 함께 추진하겠다고 밝혔다.

openai.com
#openai#privacy-design#ai-safety#llm
GPT-5.5 Instant: smarter, clearer, and more personalized
Article2026년 5월 5일

GPT-5.5 Instant: smarter, clearer, and more personalized

OpenAI는 GPT 5.5 Instant를 ChatGPT의 기본 모델로 업데이트해 더 정확하고 간결하며 개인 맥락에 맞는 답변을 제공한다고 발표했다.

openai.com
#openai#privacy-design#context-compression#prompt-library
New ways to buy ChatGPT ads
Article2026년 5월 5일

New ways to buy ChatGPT ads

OpenAI는 ChatGPT 광고 파일럿을 확대하며 파트너 구매, 베타 셀프서브 Ads Manager, CPC 입찰, 전환 측정 도구를 추가하되 답변 독립성·대화 프라이버시·사용자 통제를 핵심 원칙으로 유지한다고 밝혔다.

openai.com
#openai#privacy-design#ai-safety#change-management
Supercomputer networking to accelerate large scale AI training
Article2026년 5월 5일

Supercomputer networking to accelerate large scale AI training

OpenAI는 대규모 AI 훈련에서 GPU 간 데이터 이동 지연과 장애를 줄이기 위해 MRC라는 새로운 다중 경로 네트워킹 프로토콜을 개발해 공개했다.

openai.com
#broadcom#nvidia#openai#privacy-design
Boosting multimodal inference performance by >10% with a single Python dictionary
Article2026년 5월 4일

Boosting multimodal inference performance by >10% with a single Python dictionary

Modal은 SGLang의 멀티모달 추론 스케줄러에서 반복적인 CUDA IPC 핸들 열기 비용을 Python dict 캐시로 제거해 Qwen2.5 VL 3B Instruct 단일 H100 벤치마크에서 처리량 16.2%, 평균 지연 10% 이상 개선했다고 설명한다.

Modal
#modal#pytorch#sglang#cuda-ipc
The distillation panic - by Nathan Lambert
Article2026년 5월 4일

The distillation panic - by Nathan Lambert

이 글은 일부 연구소의 API 우회·탈옥·신원 위장 행위를 ‘증류 공격’으로 부르는 담론이 표준적인 모델 증류 기법 전체를 범죄처럼 낙인찍고, 개방형 가중치 생태계에 과잉 규제를 초래할 수 있다고 경고한다.

interconnects.ai
#anthropic#privacy-design#service-design#agent-routing
Granite Embedding Multilingual R2: Open Apache 2.0 Multilingual Embeddings with 32K Context — Best Sub-100M Retrieval Quality
Article2026년 5월 4일

Granite Embedding Multilingual R2: Open Apache 2.0 Multilingual Embeddings with 32K Context — Best Sub-100M Retrieval Quality

IBM의 Granite Embedding Multilingual R2는 200개 이상의 언어와 32K 토큰 문맥을 지원하며, 97M 소형 모델과 311M 고성능 모델로 다국어·장문·코드 검색의 성능과 배포 효율을 함께 높인 Apache 2.0 임베딩 모델군이다.

huggingface.co
#ai-architecture#multimodal#agent-memory#retrieval-index
이전1…2425262728…5426 / 54다음