Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#service-design
Tag364건YouTube 3Article 361

#service-design

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#semiconductors공동문서 286 · 연관도 43%#applications공동문서 270 · 연관도 43%#llm공동문서 297 · 연관도 42%#agent-routing공동문서 137 · 연관도 31%#ai-architecture공동문서 131 · 연관도 30%#privacy-design공동문서 109 · 연관도 27%#anthropic공동문서 93 · 연관도 22%#agent-memory공동문서 90 · 연관도 20%#agent-deployment공동문서 69 · 연관도 19%#workflow-automation공동문서 37 · 연관도 16%
AI Agents (and humans) do better with good abstractions
Article2025년 6월 3일

AI Agents (and humans) do better with good abstractions

Convex의 사례는 좋은 추상화가 개발자뿐 아니라 AI 에이전트도 복잡한 풀스택 앱을 더 안정적으로 만들게 한다는 점을 보여준다.

stack.convex.dev
#service-design#ai-architecture#agent-routing#context-compression
Blazingly fast whisper transcriptions with Inference Endpoints
Article2025년 5월 13일

Blazingly fast whisper transcriptions with Inference Endpoints

허깅페이스는 vLLM과 GPU 최적화를 적용한 새로운 Whisper 추론 엔드포인트를 공개해 기존 버전 대비 최대 8배 빠른 처리 성능을 제공하면서도 전사 품질을 유지했다.

huggingface.co
#service-design#agent-deployment#ai-infrastructure#capex-cycle
Introducing HealthBench
Article2025년 5월 12일

Introducing HealthBench

HealthBench는 262명의 의사가 설계한 5,000개의 현실적 건강 대화와 48,562개의 맞춤형 평가 기준을 통해 의료 환경에서 AI의 유용성·안전성·신뢰성을 측정하는 벤치마크다.

openai.com
#openai#privacy-design#service-design#agent-routing
Evolving OpenAI’s structure
Article2025년 5월 5일

Evolving OpenAI’s structure

OpenAI는 비영리법인의 통제와 기존 사명을 유지하면서 영리 유한책임회사를 공익기업으로 전환해 대규모 자본 조달, 인공지능의 보편적 이용, 안전한 범용인공지능 개발을 함께 추진하겠다고 밝혔다.

openai.com
#anthropic#openai#privacy-design#service-design
Lowe’s leverages AI to power home improvement retail
Article2025년 5월 5일

Lowe’s leverages AI to power home improvement retail

로우스는 인공지능을 단순한 판매 기술이 아니라 고객의 주택 개량 프로젝트를 안내하고 직원의 전문성을 확장하며 유통 운영을 개선하는 전사적 사업 전환 수단으로 활용하고 있다.

openai.com
#openai#privacy-design#service-design#agent-deployment
Accelerating Qwen3-8B Agent on Intel® Core™ Ultra with Depth-Pruned Draft Models
Article2025년 4월 30일

Accelerating Qwen3-8B Agent on Intel® Core™ Ultra with Depth-Pruned Draft Models

이 글은 Qwen3 8B를 Intel® Core™ Ultra에서 더 빠르게 실행하기 위해 OpenVINO.GenAI의 추측 디코딩과 깊이 가지치기된 Qwen3 0.6B 드래프트 모델을 결합해 약 1.4배 속도 향상을 얻은 과정을 설명한다.

huggingface.co
#service-design#agent-routing#capex-cycle#context-compression
LLM Inference on Edge: A Fun and Easy Guide to run LLMs via React Native on your Phone!
Article2025년 4월 21일

LLM Inference on Edge: A Fun and Easy Guide to run LLMs via React Native on your Phone!

이 글은 React Native 앱에서 Hugging Face의 GGUF 모델을 내려받고 llama.rn으로 로컬 실행해, Android와 iOS에서 오프라인 LLM 채팅 앱을 만드는 과정을 안내한다.

huggingface.co
#privacy-design#service-design#llm#semiconductors
Get your VLM running in 3 simple steps on Intel CPUs
Article2025년 4월 8일

Get your VLM running in 3 simple steps on Intel CPUs

소형 비전 언어 모델 SmolVLM2를 Optimum Intel과 OpenVINO로 변환·양자화·추론하면 별도 GPU 없이도 Intel CPU에서 지연시간과 처리량을 크게 개선할 수 있다.

huggingface.co
#privacy-design#service-design#multimodal#agent-memory
Zendesk uses OpenAI to build adaptive service agents focused on resolutions
Article2025년 3월 27일

Zendesk uses OpenAI to build adaptive service agents focused on resolutions

Zendesk는 OpenAI 모델을 활용해 정해진 대화 흐름에 머무르던 기존 봇을 넘어, 고객 문제의 해결을 목표로 대화를 이끌고 맥락에 맞춰 행동하는 적응형 서비스 AI 에이전트를 파일럿 운영하고 있다.

openai.com
#openai#zendesk#gpt-4o#o3-mini
Introducing next-generation audio models in the API
Article2025년 3월 20일

Introducing next-generation audio models in the API

OpenAI는 음성 에이전트 구축을 위해 더 정확한 음성 텍스트 모델과 조절 가능한 텍스트 음성 모델을 API에 출시했다.

openai.com
#service-design#multimodal#ai-safety#llm
EliseAI improves housing and healthcare efficiency with AI
Article2025년 3월 18일

EliseAI improves housing and healthcare efficiency with AI

EliseAI는 주거·의료 현장의 기존 업무 방식을 대화형 AI로 자동화하고, 고객의 사업 성과와 이용자 경험을 함께 개선하는 데 집중한다.

openai.com
#service-design#ai-architecture#agent-routing#change-management
LY Corporation: Driving growth and ‘WOW’ moments with OpenAI
Article2025년 3월 12일

LY Corporation: Driving growth and ‘WOW’ moments with OpenAI

LY 코퍼레이션은 오픈AI의 모델과 응용 프로그램 인터페이스를 사내 업무와 라인·야후! 재팬 서비스에 적용해 32개 활용 사례를 구축하고, 사용자 경험 개선과 대규모 생산성·매출 성장을 추진하고 있다.

openai.com
#openai#privacy-design#service-design#agent-routing
Detecting misbehavior in frontier reasoning models
Article2025년 3월 10일

Detecting misbehavior in frontier reasoning models

OpenAI는 프런티어 추론 모델이 보상 구조의 허점을 찾아 악용할 수 있으며, 체인오브소트(CoT)를 다른 LLM으로 감시하면 이런 의도를 잘 포착할 수 있지만 CoT 자체를 강하게 최적화하면 모델이 악의적 의도를 숨기게 된다고 보고했다.

openai.com
#openai#gpt-4o#reward-hacking#chain-of-thought
Hugging Face and FriendliAI partner to supercharge model deployment on the Hub
Article2025년 3월 4일

Hugging Face and FriendliAI partner to supercharge model deployment on the Hub

Hugging Face와 FriendliAI는 Hugging Face Hub의 “Deploy this model” 버튼 안에 FriendliAI Endpoints를 통합해 생성형 AI 모델 배포와 추론 운영을 더 빠르고 간단하게 만들었다.

huggingface.co
#nvidia#service-design#agent-deployment#agent-routing
Reinforcement Learning Heats Up, White House Orders Muscular AI Policy, and more...
Article2025년 1월 29일

Reinforcement Learning Heats Up, White House Orders Muscular AI Policy, and more...

DeepSeek R1의 공개는 중국 생성 AI의 추격, 오픈 웨이트 모델의 공급망 중요성, 기반 모델 가격 하락, 그리고 강화학습 기반 추론 모델·컴퓨터 사용 에이전트의 부상을 동시에 부각시켰다.

@DeepLearningAI
#anthropic#service-design#ai-architecture#multimodal
Welcome to Inference Providers on the Hub 🔥
Article2025년 1월 28일

Welcome to Inference Providers on the Hub 🔥

허깅페이스는 팔, 리플리케이트, 삼바노바, 투게더 에이아이의 서버리스 추론을 허브 모델 페이지와 파이썬·자바스크립트 SDK, HTTP 라우터에 통합하고, 자체 제공자 키와 허깅페이스 경유 방식 중 하나를 선택할 수 있게 했다.

huggingface.co
#privacy-design#service-design#agent-routing#ai-safety
Hugging Face models in Amazon Bedrock
Article2024년 12월 9일

Hugging Face models in Amazon Bedrock

Models mentioned in this article 1 More Articles from our Blog aws partnerships How to deploy and fine tune DeepSeek models on AWS 56 January 30, 2025…

huggingface.co
#service-design#agent-deployment#llm#semiconductors
OpenAI o1 System Card
Article2024년 12월 5일

OpenAI o1 System Card

OpenAI o1 계열은 강화학습 기반의 숙고형 추론으로 안전 정책 준수와 탈옥 공격 저항성을 높였지만, 향상된 지능이 새로운 위험을 만들 수 있어 지속적인 정렬·스트레스 테스트·위험 관리가 필요하다.

openai.com
#openai#privacy-design#service-design#multimodal
Welcome PaliGemma 2 – New vision language models by Google
Article2024년 12월 5일

Welcome PaliGemma 2 – New vision language models by Google

PaliGemma 2는 SigLIP 이미지 인코더와 Gemma 2 언어 모델을 결합하고, 3B·10B·28B 규모와 세 가지 해상도, 간편한 미세조정 및 양자화 지원을 제공하는 구글의 새로운 오픈 비전 언어 모델 제품군이다.

huggingface.co
#hotel-review#service-design#multimodal#search-advertising
An Object Sync Engine for Local-first Apps
Article2024년 11월 13일

An Object Sync Engine for Local-first Apps

이 글은 로컬 퍼스트 웹 앱에 적합한 객체 동기화 엔진을 정의하고, 로컬 저장소·서버 저장소·동기화 프로토콜의 역할을 기존 사례와 컨벡스의 방향으로 비교한다.

stack.convex.dev
#service-design#ai-architecture#agent-memory#retrieval-index
Rearchitecting Hugging Face Uploads and Downloads
Article2024년 9월 27일

Rearchitecting Hugging Face Uploads and Downloads

허깅페이스는 대용량 모델·데이터셋 전송의 한계를 극복하기 위해 바이트 단위 중복 제거와 검증을 수행하는 콘텐츠 주소 기반 저장소(CAS)를 도입하고, 전 세계 업로드 분포에 맞춰 세 지역에 노드를 배치하는 새로운 업로드·다운로드 구조를 설계했다.

huggingface.co
#service-design#ai-architecture#llm#semiconductors
Mini-R1: Reproduce Deepseek R1 „aha moment“ a RL tutorial
Article2024년 9월 25일

Mini-R1: Reproduce Deepseek R1 „aha moment“ a RL tutorial

3B 규모의 큐원 모델을 카운트다운 산술 과제와 그룹 상대 정책 최적화로 훈련해, 별도의 인간 피드백 없이 풀이를 검토하고 탐색하는 딥시크 R1의 작은 ‘아하 순간’을 재현한 강화학습 실험이다.

huggingface.co
#nvidia#service-design#agent-memory#capex-cycle
Expert Support case study: Bolstering a RAG app with LLM-as-a-Judge
Article2024년 9월 13일

Expert Support case study: Bolstering a RAG app with LLM-as-a-Judge

Digital Green은 농업 RAG 챗봇 Farmer.chat의 응답 품질을 대규모로 평가하기 위해 LLM as a Judge 체계를 구축하고, 충실도와 답변률의 균형을 기준으로 생성 모델을 비교·선정했다.

huggingface.co
#service-design#ai-architecture#agent-memory#context-compression
Handling 300k requests per day: an adventure in scaling
Article2024년 9월 13일

Handling 300k requests per day: an adventure in scaling

Firecrawl 팀은 사용자 증가로 API와 크롤링 작업이 불안정해지자 큐 락 설정, 스크레이프 처리 구조, 크롤 작업 단위, 큐 라이브러리, Redis 인프라를 단계적으로 바꾸며 하루 30만 요청 규모에 대응했다.

firecrawl.dev
#privacy-design#service-design#ai-architecture#agent-routing
이전1…1314151615 / 16다음