Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#ai-architecture
Tag508건YouTube 42Article 466

#ai-architecture

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#llm공동문서 421 · 연관도 51%#applications공동문서 315 · 연관도 42%#semiconductors공동문서 295 · 연관도 37%#agent-routing공동문서 190 · 연관도 36%#service-design공동문서 131 · 연관도 30%#agent-memory공동문서 160 · 연관도 30%#multimodal공동문서 85 · 연관도 26%#vision-language-models공동문서 82 · 연관도 26%#retrieval-index공동문서 87 · 연관도 24%#gpu공동문서 48 · 연관도 24%
GPT-5.5 Outperforms (and Hallucinates), Kimi K2.6 Leads Open LLMs, AI Strains Climate Pledges, and more...
Article2026년 5월 1일

GPT-5.5 Outperforms (and Hallucinates), Kimi K2.6 Leads Open LLMs, AI Strains Climate Pledges, and more...

글은 최신 AI 모델 사용법 교육 안내를 시작으로, GPT 5.5의 높은 벤치마크 성과와 환각 문제, 대형 AI 기업의 데이터센터 확장으로 흔들리는 탄소 감축 약속, 그리고 오픈 가중치 모델 Kimi K2.6의 경쟁력을 함께 다룬다.

@DeepLearningAI
#kimi-k2-5#anthropic#agent-swarms#service-design
Granite 4.0 3B Vision: Compact Multimodal Intelligence for Enterprise Documents
Article2026년 4월 30일

Granite 4.0 3B Vision: Compact Multimodal Intelligence for Enterprise Documents

Granite 4.0 3B Vision은 복잡한 기업 문서의 표·차트·핵심 값 쌍을 정밀하게 추출하도록 설계된 30억 매개변수급 소형 비전·언어 모델이다.

huggingface.co
#ai-architecture#multimodal#agent-routing#workflow-automation
Demis Hassabis: We''''re Three Quarters of the Way to AGI
YouTube2026년 4월 30일

Demis Hassabis: We''''re Three Quarters of the Way to AGI

데미스 하사비스는 AGI를 과학·의학·세계 이해를 확장하는 “정밀한 도구”로 보고, 2030년 전후 AGI 가능성과 AI for Science의 폭발적 확장을 주장합니다.

Sequoia Capital
#agi#demis#change-management#organizational-redesign
Transcript: ‘How Stripe Is Building for an Agent-native World’
Article2026년 4월 29일

Transcript: ‘How Stripe Is Building for an Agent-native World’

Stripe의 Emily Glassberg Sands는 인터넷 경제의 주체가 인간에서 AI 에이전트와 소프트웨어로 확장되면서 결제, 과금, 사기 탐지, 신원 인프라가 거래 순간이 아니라 고객 생애주기 전체를 다루도록 바뀌고 있다고 설명한다.

Dan Shipper
#anthropic#privacy-design#service-design#ai-architecture
What to Learn, Build, and Skip in AI Agents (2026)
Article2026년 4월 29일

What to Learn, Build, and Skip in AI Agents (2026)

AI 에이전트 시대에는 매주 바뀌는 프레임워크보다 오래 남는 기본 원리, 평가 체계, 도구 설계, 안전한 실행 환경을 고르는 능력이 더 중요하다.

Rohit
#braintrust#langfuse#langgraph#langsmith
The Economics of the Last 1 cm - Five Companies Standing at the Edge of AI GPU Power
Article2026년 4월 28일

The Economics of the Last 1 cm - Five Companies Standing at the Edge of AI GPU Power

AI GPU 전력 공급의 병목은 랙이나 보드가 아니라 GPU 다이 직전의 “마지막 1cm”에 있으며, 이 구간을 둘러싼 기업들의 수익화 방식이 서로 다르게 갈라지고 있다.

Nutty
#ai-architecture#nvidia#openvreg#ai-gpu
GLM 5.1 Thinks Strategically, Data-Center Revolt Intensifies, When Helpful LLMs Turn Unhelpful, and more...
Article2026년 4월 24일

GLM 5.1 Thinks Strategically, Data-Center Revolt Intensifies, When Helpful LLMs Turn Unhelpful, and more...

본문은 코딩 에이전트가 프론트엔드에는 큰 가속을 주지만 백엔드·인프라·연구로 갈수록 한계가 커진다는 판단과, 장시간 자율 작업을 지향하는 GLM 5.1 및 초기 산업 현장의 휴머노이드 로봇 배치를 다룬다.

@DeepLearningAI
#anthropic#ai-architecture#multimodal#agent-routing
Every AI trend you need to know in 2026
Article2026년 4월 22일

Every AI trend you need to know in 2026

2026년 AI의 핵심 변화는 더 좋은 프롬프트를 쓰는 것이 아니라, 지속되는 맥락·도구·메모리를 갖춘 작업 환경을 AI 주변에 구축하는 방향으로 이동했다는 주장입니다.

Rohit
#openclaw#cursor#claude-code#copilot-workspace
Speeding up agentic workflows with WebSockets in the Responses API
Article2026년 4월 22일

Speeding up agentic workflows with WebSockets in the Responses API

Responses API의 WebSocket 모드는 Codex식 에이전트 루프에서 반복 API 요청과 상태 재처리 비용을 줄여, 더 빨라진 추론 속도를 실제 사용자 체감 속도로 전달하도록 만든 개선이다.

openai.com
#codex#openai#websocket#responses-api
Transcript: ‘The AI Sandwich: Where Humans Excel in an AI World’
Article2026년 4월 22일

Transcript: ‘The AI Sandwich: Where Humans Excel in an AI World’

이 대담은 AI가 실행의 ‘중간’을 맡고 인간이 문제 설정과 최종 판단의 ‘처음과 끝’을 책임지는 ‘AI 샌드위치’ 모델로, 에이전트 시대에도 인간의 역할이 어디에 남는지를 설명한다.

Dan Shipper
#privacy-design#ai-architecture#agent-routing#llm
Moving past bots vs. humans
Article2026년 4월 21일

Moving past bots vs. humans

원문은 웹 보호의 핵심 기준이 더 이상 ‘봇인가 인간인가’가 아니라, 요청 주체의 의도·행동·책임성·가치 반환 여부가 되어야 한다고 설명한다.

blog.cloudflare.com
#privacy-design#service-design#ai-architecture#search-advertising
Orchestrating AI Code Review at scale
Article2026년 4월 20일

Orchestrating AI Code Review at scale

Cloudflare는 대규모 코드 리뷰 병목을 줄이기 위해 OpenCode 기반의 CI 네이티브 오케스트레이션 시스템을 만들고, 여러 전문 AI 리뷰어와 조정자 에이전트를 조합해 구조화된 리뷰를 수행하도록 했다.

The Cloudflare Blog
#cloudflare#configurecontext#gitlab#opencode
The AI engineering stack we built internally — on the platform we ship
Article2026년 4월 20일

The AI engineering stack we built internally — on the platform we ship

Cloudflare는 자체 플랫폼 위에 인증, LLM 라우팅, 추론, MCP 포털, 설정 자동화, 사용량 추적을 결합한 내부 AI 엔지니어링 스택을 구축했고, 최근 30일 동안 R&D 조직의 93%가 이를 활용했다.

blog.cloudflare.com
#kimi-k2-5#anthropic#service-design#ai-architecture
Agents that remember: introducing Agent Memory
Article2026년 4월 17일

Agents that remember: introducing Agent Memory

Cloudflare는 에이전트 대화에서 중요한 정보를 추출·검증·분류해 필요할 때만 검색해 주는 관리형 서비스 Agent Memory의 비공개 베타를 소개하며, 긴 컨텍스트 창만으로 해결되지 않는 컨텍스트 품질 문제를 겨냥한다.

Cloudflare
#cloudflare#agent-memory#agents-sdk#cloudflare-workers
Introducing the Agent Readiness score. Check to see if your site is agent-ready
Article2026년 4월 17일

Introducing the Agent Readiness score. Check to see if your site is agent-ready

Cloudflare는 웹사이트가 AI 에이전트와 얼마나 잘 상호작용할 수 있는지 평가하는 Agent Readiness 점수와 isitagentready.com, 그리고 관련 표준 채택 현황을 추적하는 Radar 데이터셋을 공개했다.

blog.cloudflare.com
#kimi-k2-5#service-design#ai-architecture#agent-deployment
Meta Pivots From Open Weights, Big Pharma Bets On AI, Regulatory Patchwork, and more...
Article2026년 4월 17일

Meta Pivots From Open Weights, Big Pharma Bets On AI, Regulatory Patchwork, and more...

원문은 AI 네이티브 소프트웨어 팀에서 역할과 병목이 재편되는 흐름을 설명한 뒤, Meta의 폐쇄형 멀티모달 모델 Muse Spark와 Eli Lilly·Insilico의 AI 신약개발 협력을 다룬다.

@DeepLearningAI
#kimi-k2-5#anthropic#meta-ai#agent-swarms
Redirects for AI Training enforces canonical content
Article2026년 4월 17일

Redirects for AI Training enforces canonical content

Cloudflare는 AI 학습 크롤러가 오래된 문서를 그대로 학습하지 않도록 기존 canonical 태그를 검증된 AI 학습 크롤러 대상 HTTP 301 리디렉션으로 강제하는 Redirects for AI Training을 출시했다.

blog.cloudflare.com
#anthropic#ai-architecture#ai-distribution#search-advertising
Shared Dictionaries: compression that keeps up with the agentic web
Article2026년 4월 17일

Shared Dictionaries: compression that keeps up with the agentic web

Cloudflare는 잦은 배포와 에이전트 트래픽으로 기존 캐싱이 약해지는 웹에서, 브라우저가 이미 가진 이전 리소스를 사전처럼 활용해 변경분만 전송하는 공유 딕셔너리 압축을 엣지에서 지원하려 한다.

@ackriv
#ai-architecture#agent-deployment#llm#semiconductors
Unweight: how we compressed an LLM 22% without sacrificing quality
Article2026년 4월 17일

Unweight: how we compressed an LLM 22% without sacrificing quality

Unweight는 LLM 가중치의 BF16 지수 바이트를 무손실로 압축하고 GPU 온칩 메모리에서 바로 복원해, 출력 품질을 유지하면서 모델 크기와 HBM 메모리 대역폭 부담을 줄이는 추론용 압축 시스템이다.

blog.cloudflare.com
#model-scaling#ai-architecture#agent-memory#context-compression
Building the foundation for running extra-large language models
Article2026년 4월 16일

Building the foundation for running extra-large language models

Cloudflare는 Workers AI에서 Kimi K2.5 같은 초대형 오픈소스 언어 모델을 빠르고 효율적으로 운영하기 위해 하드웨어 구성, prefill/decode 분리, 프롬프트 캐싱, KV 캐시 최적화, speculative decoding, 자체 추론 엔진 Infire를 조합하고 있다고 설명한다.

@_mchenco
#kimi-k2-5#nvidia#ai-architecture#agent-memory
The PR you would have opened yourself
Article2026년 4월 16일

The PR you would have opened yourself

transformers 모델을 mlx lm으로 빠르고 정확하게 이식하도록 돕는 스킬과 비에이전트 테스트 하네스를 구축하되, 코드 소유권과 최종 판단은 기여자와 리뷰어에게 남겨 두는 접근을 설명한다.

huggingface.co
#privacy-design#ai-architecture#multimodal#llm
Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers
Article2026년 4월 16일

Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers

이 글은 Sentence Transformers로 텍스트와 이미지 등 여러 모달리티를 다루는 임베딩 모델을 자체 데이터에 맞게 파인튜닝하는 방법을, 시각 문서 검색 사례를 중심으로 설명합니다.

huggingface.co
#ai-architecture#multimodal#llm#vision-language-models
State of Open Source on Hugging Face: Spring 2026
Article2026년 4월 14일

State of Open Source on Hugging Face: Spring 2026

허깅페이스의 오픈소스 인공지능 생태계는 사용자·모델·데이터셋과 파생 창작물이 급증한 가운데, 중국과 독립 개발자의 영향력 확대, 소형 모델 중심의 실용적 채택, 국가 주권과 지역 생태계의 부상을 동시에 보여준다.

huggingface.co
#ai-architecture#multimodal#agent-deployment#agent-routing
Anthropic’s Claude Mythos Problem, Dark DNA Unveiled, Pitfalls for Assistive Models, and more...
Article2026년 4월 10일

Anthropic’s Claude Mythos Problem, Dark DNA Unveiled, Pitfalls for Assistive Models, and more...

이 글은 AI 코딩 에이전트가 소프트웨어 엔지니어링과 고용 논의를 바꾸는 흐름을 짚고, Anthropic의 Claude Mythos Preview가 제기한 사이버보안 위험과 시각장애인을 위한 보조 AI의 심리적·사회적 함의를 함께 다룬다.

deeplearning.ai
#anthropic#ai-architecture#multimodal#agent-routing
이전1…1112131415…2213 / 22다음