Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#ai-architecture
Tag508건YouTube 42Article 466

#ai-architecture

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#llm공동문서 421 · 연관도 51%#applications공동문서 315 · 연관도 42%#semiconductors공동문서 295 · 연관도 37%#agent-routing공동문서 190 · 연관도 36%#service-design공동문서 131 · 연관도 30%#agent-memory공동문서 160 · 연관도 30%#multimodal공동문서 85 · 연관도 26%#vision-language-models공동문서 82 · 연관도 26%#retrieval-index공동문서 87 · 연관도 24%#gpu공동문서 48 · 연관도 24%
AI 타고 날아갈 줄 알았던 SaaS 회사들의 좌절, 누가누가 떨고있나?
YouTube2026년 3월 3일

AI 타고 날아갈 줄 알았던 SaaS 회사들의 좌절, 누가누가 떨고있나?

AI 에이전트가 무너뜨리는 것은 개별 기능이 아니라 사람 수에 기대던 SaaS의 워크플로우와 인당 과금 구조이며, 앞으로의 승자는 기능 앱보다 데이터 저장소·오케스트레이션·전환비용을 장악한 플랫폼일 가능성이 높다. 투자 판단의 핵심은 AI 도입 여부가 아니라 에이전트 시대에도 가격 결정권과 락인을 유지할 수 있는지다.

티타임즈TV
#openclaw#arr#saas#agentic-ai
가스터빈을 지금 봐야만 하는 매우 중요한 이유ㅣ김효식 삼성액티브자산운용 팀장
YouTube2026년 3월 3일

가스터빈을 지금 봐야만 하는 매우 중요한 이유ㅣ김효식 삼성액티브자산운용 팀장

미국 전력난의 단기 해법은 SMR 같은 장기 옵션보다 가스터빈·재생에너지·ESS·연료전지의 현실 조합에 있고, 투자 포인트도 업종 전체가 아니라 미국의 비중국 공급망 재편에 직접 연결된 한국 ESS·연료전지 밸류체인 선별로 좁혀진다.

이효석아카데미
#openclaw#ess#feoc#smr
The design process is dead. Here''''s what''''s replacing it.
YouTube2026년 3월 1일

The design process is dead. Here''''s what''''s replacing it.

AI가 구현 속도를 폭발적으로 끌어올린 시대에는 디자인의 중심이 정교한 선행 목업 제작에서, 빠르게 만들어지는 결과물을 정렬하고 다듬고 방향을 부여하며 책임 있게 판단하는 일로 이동하고 있다.

Lenny''s Podcast
#openclaw#anthropic#jenny-wen#ai-tools
Gemini Seizes the Lead, Investors Panic Over Agentic AI, Optimism at Global AI Summit, and more...
Article2026년 2월 27일

Gemini Seizes the Lead, Investors Panic Over Agentic AI, Optimism at Global AI Summit, and more...

원문은 DeepLearning.AI의 Skill Builder 소개로 시작해 Gemini 3.1 Pro Preview의 벤치마크 우위, 뉴델리 AI Impact Summit의 낙관적 기조, 그리고 에이전트형 AI가 SaaS 기업 가치에 준 충격을 다룹니다.

@DeepLearningAI
#anthropic#privacy-design#service-design#ai-architecture
Knowledge Priming
Article2026년 2월 24일

Knowledge Priming

AI 코딩 어시스턴트가 훈련 데이터의 평균적인 패턴으로 흘러가지 않도록, 프로젝트 맥락을 버전 관리되는 프라이밍 문서로 제공해야 한다는 글이다.

martinfowler.com
#service-design#ai-architecture#agent-memory#retrieval-index
The Convex Marketing Journey 2022-2026
Article2026년 2월 23일

The Convex Marketing Journey 2022-2026

Convex는 2022년부터 2026년까지 여러 메시지와 태그라인을 실험하며, 제품의 본질을 설명하는 일보다 ‘누구에게 어떤 의미로 들리는가’를 맞추는 일이 더 중요하다는 교훈을 얻었다.

stack.convex.dev
#ai-architecture#agent-routing#semiconductors#applications
Our Multi-Agent Architecture for Smarter Advertising
Article2026년 2월 19일

Our Multi-Agent Architecture for Smarter Advertising

Spotify Engineering은 광고 구매 채널마다 흩어진 의사결정 로직을 통합하기 위해, 기존 Ads API를 도구처럼 호출하는 멀티 에이전트 기반 미디어 플래닝 아키텍처를 도입했다고 설명한다.

Spotify Engineering
#spotify#spotify-engineering#spotify-ads-api#spotify-ads-manager
Diffusers welcomes FLUX-2
Article2026년 2월 17일

Diffusers welcomes FLUX-2

FLUX.2는 처음부터 새로 사전학습한 이미지 생성·편집 모델로, Diffusers는 새로운 구조와 다중 이미지 참조 기능부터 8~80GB급 그래픽 메모리 환경별 추론 방법까지 구체적으로 소개한다.

huggingface.co
#ai-architecture#multimodal#agent-memory#agent-routing
Beyond rate limits: scaling access to Codex and Sora
Article2026년 2월 13일

Beyond rate limits: scaling access to Codex and Sora

OpenAI는 Codex와 Sora의 빠른 사용 증가로 기존 속도 제한만으로는 사용자 흐름을 지키기 어렵다고 보고, 속도 제한과 크레딧을 실시간으로 결합하는 자체 접근·사용량·잔액 시스템을 구축했다.

openai.com
#openai#privacy-design#service-design#ai-architecture
Training Design for Text-to-Image Models: Lessons from Ablations
Article2026년 2월 13일

Training Design for Text-to-Image Models: Lessons from Ablations

Photoroom 팀은 PRX 1.2B 텍스트 이미지 모델을 기준점으로 삼아, 표현 정렬과 보조 목적 함수가 훈련 수렴·품질·처리량에 어떤 영향을 주는지 절제된 ablation 방식으로 검증했다.

huggingface.co
#muon-optimizer#ai-architecture#agent-routing#llm
Preserving the Freedom to Learn in AI
Article2026년 2월 11일

Preserving the Freedom to Learn in AI

이 글은 웹이 발전해 온 ‘합법적으로 접근 가능한 정보를 읽고, 분석하고, 그 위에 구축할 자유’를 AI 시대에도 지켜야 하며, 출판자의 통제권과 이용자의 학습 자유를 시장·기술 표준·공공정책으로 균형 있게 조정해야 한다고 주장한다.

Matt Perault
#privacy-design#service-design#ai-architecture#ai-ai-ai
How to connect Convex to RunPod for serverless GPU workloads
Article2026년 2월 9일

How to connect Convex to RunPod for serverless GPU workloads

이 글은 Convex가 파일 업로드 직후 RunPod 서버리스 GPU 워커를 호출하고, 워커가 처리 결과를 다시 Convex Storage와 데이터베이스에 반영하는 배경 제거 워크플로를 설명한다.

stack.convex.dev
#ai-architecture#agent-memory#agent-routing#capex-cycle
How AI tools can redefine universal design to increase accessibility
Article2026년 2월 5일

How AI tools can redefine universal design to increase accessibility

Google Research는 장애 커뮤니티와의 공동 설계를 바탕으로, 사용자의 맥락과 필요에 맞춰 스스로 조정되는 멀티모달 AI 기반 ‘Natively Adaptive Interfaces’가 보편적 디자인과 접근성을 새롭게 정의할 수 있다고 설명한다.

research.google
#privacy-design#ai-architecture#multimodal#agent-routing
Unlocking the Codex harness: how we built the App Server
Article2026년 2월 4일

Unlocking the Codex harness: how we built the App Server

OpenAI는 여러 Codex 제품이 같은 에이전트 루프와 실행 로직을 재사용할 수 있도록, Codex harness를 양방향 JSON RPC 기반 App Server로 노출하는 구조를 만들었다.

openai.com
#openai#privacy-design#ai-architecture#agent-routing
Why Nvidia builds open models with Bryan Catanzaro
Article2026년 2월 4일

Why Nvidia builds open models with Bryan Catanzaro

엔비디아는 AI를 개방형 인프라로 확산시키는 동시에 모델 연구에서 얻은 지식을 가속 컴퓨팅 제품에 반영하기 위해 네모트론 모델·데이터·도구를 공개하고 있다.

Nathan Lambert
#nemotron#nvidia#megatron-lm#reproducible-model-development
Collaborating on a nationwide randomized study of AI in real-world virtual care
Article2026년 2월 3일

Collaborating on a nationwide randomized study of AI in real-world virtual care

Google은 Included Health와 함께 실제 가상진료 환경에서 대화형 의료 AI의 안전성·유용성·한계를 평가하기 위한 전국 단위 무작위 연구를 추진한다.

research.google
#ai-architecture#multimodal#agent-routing#llm
AprielGuard: A Guardrail for Safety and Adversarial Robustness in Modern LLM Systems
Article2026년 1월 27일

AprielGuard: A Guardrail for Safety and Adversarial Robustness in Modern LLM Systems

AprielGuard는 현대 LLM의 안전 위험과 프롬프트 주입·탈옥·메모리 오염 등 적대적 공격을 통합적으로 감지하도록 설계된 8B 안전·보안 가드레일 모델이다.

huggingface.co
#privacy-design#ai-architecture#agent-memory#agent-routing
How Credal Extracts 6M+ URLs Monthly to Power Production AI Agents
Article2026년 1월 26일

How Credal Extracts 6M+ URLs Monthly to Power Production AI Agents

Credal은 Firecrawl을 활용해 월 600만 개 이상의 URL을 스크래핑·크롤링하고, 기업용 AI 에이전트가 최신 외부 정보와 문서 기반 지식을 안전하게 활용하도록 지원한다.

Eric Ciarla
#credal#firecrawl#jack-fischer#url-extraction-volume
Unrolling the Codex agent loop
Article2026년 1월 23일

Unrolling the Codex agent loop

OpenAI는 Codex CLI의 핵심인 에이전트 루프가 사용자 입력, 모델 추론, 도구 호출, 실행 결과 반영을 반복하며 소프트웨어 작업을 완수하는 방식을 설명한다.

openai.com
#openai#codex-cli#codex-cloud#responses-api
Inside Praktika's conversational approach to language learning
Article2026년 1월 22일

Inside Praktika's conversational approach to language learning

프락티카는 개인화된 다중 에이전트, 적시 기억 검색, 비원어민 음성 인식을 결합해 실제 상황에서 자신 있게 말하는 능력을 키우는 대화형 언어 학습 앱이다.

openai.com
#ai-architecture#agent-memory#retrieval-index#ai-9-ai
Scaling PostgreSQL to power 800 million ChatGPT users
Article2026년 1월 22일

Scaling PostgreSQL to power 800 million ChatGPT users

OpenAI는 ChatGPT와 API의 8억 사용자 규모 트래픽을 감당하기 위해 단일 primary PostgreSQL과 전 세계 약 50개 읽기 replica를 기반으로 읽기 중심 워크로드를 수백만 QPS까지 확장했다.

openai.com
#openai#privacy-design#service-design#ai-architecture
Differential Transformer V2
Article2026년 1월 20일

Differential Transformer V2

DIFF V2는 같은 GQA 그룹의 두 쿼리 헤드 출력을 토큰·헤드별 계수로 차감해, 키·값 헤드와 표준 어텐션 커널을 그대로 유지하면서 디코딩 효율, 학습 안정성, 표현 자유도, 출력 투영의 매개변수 효율을 함께 개선한 구조다.

huggingface.co
#ai-architecture#multimodal#llm#semiconductors
Unlocking health insights: Estimating advanced walking metrics with smartwatches
Article2026년 1월 15일

Unlocking health insights: Estimating advanced walking metrics with smartwatches

구글 연구진은 대규모 검증 연구를 통해 손목 착용 스마트워치가 보행 속도, 보폭, 지지 시간 등 고급 보행 지표를 스마트폰과 비슷한 수준으로 정확하고 신뢰성 있게 추정할 수 있음을 보였다.

research.google
#ai-architecture#llm#semiconductors#applications
Dynamic surface codes open new avenues for quantum error correction
Article2026년 1월 13일

Dynamic surface codes open new avenues for quantum error correction

구글 퀀텀 AI는 Willow 프로세서에서 동적 표면 코드를 실험해 결합기 수를 줄이고, 누설로 인한 상관 오류를 억제하며, iSWAP 게이트 기반 오류 정정 가능성을 확인했다.

research.google
#privacy-design#ai-architecture#llm#semiconductors
이전1…1415161718…2216 / 22다음