Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#semiconductors
Tag1216건YouTube 36Article 1180

#semiconductors

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#llm공동문서 1063 · 연관도 83%#applications공동문서 960 · 연관도 83%#agent-routing공동문서 483 · 연관도 59%#agent-memory공동문서 468 · 연관도 57%#privacy-design공동문서 385 · 연관도 51%#context-compression공동문서 360 · 연관도 51%#agent-deployment공동문서 302 · 연관도 46%#service-design공동문서 286 · 연관도 43%#retrieval-index공동문서 211 · 연관도 38%#ai-architecture공동문서 295 · 연관도 38%
Accelerate ND-Parallel: A guide to Efficient Multi-GPU Training
Article2025년 8월 12일

Accelerate ND-Parallel: A guide to Efficient Multi-GPU Training

Accelerate와 Axolotl은 데이터·완전 샤딩 데이터·텐서·컨텍스트 병렬화를 조합해 대규모 모델의 메모리 사용량, 계산량, 장치 간 통신 비용을 균형 있게 설계하도록 지원한다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#capex-cycle
Basis scales accounting by turning OpenAI model progress into trusted agents
Article2025년 8월 12일

Basis scales accounting by turning OpenAI model progress into trusted agents

베이시스는 업무별로 적합한 오픈AI 모델을 배치하고 판단 근거를 검토 가능하게 공개함으로써, 회계 자동화를 신뢰할 수 있는 업무 위임 체계로 확장하고 있다.

openai.com
#openai#ai-architecture#agent-routing#workflow-automation
Enabling physician-centered oversight for AMIE
Article2025년 8월 12일

Enabling physician-centered oversight for AMIE

g AMIE는 환자 문진과 의료 기록 초안을 맡되 개별화된 의학적 조언은 금지하고, 최종 판단과 환자 전달은 감독 의사가 검토·수정하도록 설계된 AMIE의 의사 중심 감독 연구 프레임워크다.

research.google
#agent-routing#llm#semiconductors#applications
OpenAI’s letter to Governor Newsom on harmonized regulation
Article2025년 8월 12일

OpenAI’s letter to Governor Newsom on harmonized regulation

OpenAI는 캘리포니아가 주별 AI 규제의 중복과 불일치를 줄이고, 연방·국제 안전 기준과 조화된 국가적 모델을 이끌어야 한다고 Newsom 주지사에게 촉구했다.

openai.com
#openai#privacy-design#ai-safety#llm
Achieving 10,000x training data reduction with high-fidelity labels
Article2025년 8월 7일

Achieving 10,000x training data reduction with high-fidelity labels

Google Ads 연구진은 광고 안전성 분류에서 전문가 고품질 라벨과 능동학습 기반 큐레이션을 결합해 LLM 미세조정에 필요한 학습 데이터를 100,000건에서 500건 미만으로 줄이면서 인간 전문가와의 정렬도를 높였다고 설명한다.

research.google
#privacy-design#ai-distribution#search-advertising#zero-click-search
Introducing GPT‑5 for developers
Article2025년 8월 7일

Introducing GPT‑5 for developers

GPT‑5는 코딩 정확도와 도구 호출, 지시 이행, 장문 맥락 검색을 함께 개선하고 개발자가 응답 방식과 추론 수준을 조절할 수 있게 설계된 OpenAI의 API용 추론 모델이다.

openai.com
#openai#long-context#privacy-design#service-design
Introducing GPT-5
Article2025년 8월 7일

Introducing GPT-5

GPT‑5는 빠른 응답과 심층 추론을 실시간으로 조합해 코딩·수학·글쓰기·건강·멀티모달 과제의 성능을 높이고, 환각과 과도한 확신을 줄인 OpenAI의 통합형 AI 시스템이다.

openai.com
#openai#privacy-design#multimodal#ai-safety
gpt-oss-120b & gpt-oss-20b Model Card
Article2025년 8월 5일

gpt-oss-120b & gpt-oss-20b Model Card

OpenAI는 gpt oss 120b와 gpt oss 20b를 공개 가중치 추론 모델로 소개하며, 에이전트형 워크플로와 도구 사용을 지원하되 공개 모델 특유의 안전 위험과 평가 결과를 함께 제시했다.

openai.com
#openai#privacy-design#agent-deployment#agent-routing
Open weights and AI for all
Article2025년 8월 5일

Open weights and AI for all

OpenAI는 고도화된 오픈 웨이트 추론 모델을 공개해 개인·비영리단체·기업·정부가 자체 인프라에서 인공지능을 실행하고 맞춤화하도록 지원하며, 이를 민주적 가치에 기반한 인공지능 생태계의 확산과 연결한다.

openai.com
#openai#privacy-design#service-design#ai-safety
How hard is it to migrate AWAY from Convex?
Article2025년 8월 4일

How hard is it to migrate AWAY from Convex?

저자는 Convex에서 벗어나는 난이도를 직접 검증하기 위해 간단한 Tanstack Start 예제를 기준으로 함수, 클라이언트 호출, 반응형 쿼리, 데이터베이스 이전을 차례로 Tanstack Start와 Drizzle/Postgres로 옮겨 보며 실제 작업량과 복잡도 변화를 살펴본다.

stack.convex.dev
#service-design#ai-architecture#agent-memory#agent-routing
Measuring Open-Source Llama Nemotron Models on DeepResearch Bench
Article2025년 8월 4일

Measuring Open-Source Llama Nemotron Models on DeepResearch Bench

NVIDIA의 개방형 딥리서치 에이전트 AI Q는 두 Llama 계열 모델과 검색·오케스트레이션 도구를 결합해 2025년 8월 DeepResearch Bench의 ‘LLM with Search’ 부문에서 종합 40.52점을 기록했다.

huggingface.co
#privacy-design#ai-architecture#search-advertising#workflow-automation
Implementing MCP Servers in Python: An AI Shopping Assistant with Gradio
Article2025년 7월 31일

Implementing MCP Servers in Python: An AI Shopping Assistant with Gradio

Gradio의 MCP 통합을 이용해 웹 탐색 도구와 IDM VTON 가상 피팅 모델을 연결하고, VS Code AI 채팅에서 사용할 수 있는 개인용 AI 쇼핑 도우미를 구현하는 방법을 설명한다.

huggingface.co
#llm#semiconductors#applications#agent-memory
Introducing AI Sheets: a tool to work with datasets using open AI models!
Article2025년 7월 31일

Introducing AI Sheets: a tool to work with datasets using open AI models!

Hugging Face AI Sheets는 스프레드시트형 인터페이스에서 프롬프트와 공개 AI 모델을 활용해 데이터셋을 만들고, 변환하고, 보강하고, 평가할 수 있게 해주는 오픈소스 노코드 도구다.

huggingface.co
#privacy-design#context-compression#prompt-library#llm
Introducing Stargate Norway
Article2025년 7월 31일

Introducing Stargate Norway

OpenAI는 노르웨이 나르비크에 재생에너지 기반의 대규모 AI 데이터센터를 조성해 유럽의 컴퓨팅 역량과 지역 AI 생태계를 확대하는 ‘스타게이트 노르웨이’ 계획을 발표했다.

openai.com
#nvidia#openai#privacy-design#ai-infrastructure
Lessons from Building an AI App Builder on Convex
Article2025년 7월 31일

Lessons from Building an AI App Builder on Convex

Chef의 사례는 AI 코딩 에이전트가 잘 작동하려면 좋은 추상화, 제한된 선택지, 타입 안전한 피드백 루프, 검증 가능한 평가 체계가 필요하다는 점을 보여준다.

stack.convex.dev
#ai-architecture#agent-routing#context-compression#prompt-library
Intercom's three lessons for creating a sustainable AI advantage
Article2025년 7월 30일

Intercom's three lessons for creating a sustainable AI advantage

인터콤은 조기 실험으로 모델 이해도를 높이고, 엄격한 평가로 도입 속도를 끌어올리며, 모델 교체가 쉬운 유연한 아키텍처를 구축해 지속 가능한 인공지능 경쟁력을 만들었다.

openai.com
#openai#privacy-design#service-design#ai-architecture
Benchmarking Language Model Performance on 5th Gen Xeon at GCP
Article2025년 7월 29일

Benchmarking Language Model Performance on 5th Gen Xeon at GCP

Google Cloud의 5세대 Xeon 기반 C4 인스턴스는 3세대 Xeon 기반 N2보다 텍스트 임베딩과 텍스트 생성에서 큰 처리량 및 비용 효율 우위를 보이며, 경량 에이전틱 AI를 CPU만으로 배포할 가능성을 보여준다.

huggingface.co
#llm#semiconductors#applications#long-context
Prefill and Decode for Concurrent Requests - Optimizing LLM Performance
Article2025년 7월 28일

Prefill and Decode for Concurrent Requests - Optimizing LLM Performance

동시 요청을 처리하는 LLM 추론에서는 프리필과 디코드의 계산 특성이 달라, 첫 토큰 지연·토큰 생성 지연·총처리량·GPU 활용률 사이의 균형에 맞춰 배칭과 프리필 청크 크기를 조정해야 한다.

huggingface.co
#agent-memory#agent-routing#retrieval-index#workflow-automation
SensorLM: Learning the language of wearable sensors
Article2025년 7월 28일

SensorLM: Learning the language of wearable sensors

센서엘엠은 5,970만 시간의 웨어러블 센서 데이터와 자동 생성 설명문을 학습해 신체 신호를 자연어로 해석하고 검색·분류·설명하는 센서 언어 기반 모델군이다.

research.google
#ai-architecture#multimodal#llm#semiconductors
4M Models Scanned: Protect AI + Hugging Face 6 Months In
Article2025년 7월 27일

4M Models Scanned: Protect AI + Hugging Face 6 Months In

Protect AI와 Hugging Face의 6개월 파트너십은 Guardian 스캔으로 447만 개 모델 버전을 검사하고 35.2만 건의 unsafe/suspicious 이슈를 찾아, 공개 모델 사용자가 보안 위험을 더 명확히 판단하도록 돕고 있다.

huggingface.co
#service-design#llm#semiconductors#applications
Ulysses Sequence Parallelism: Training with Million-Token Contexts
Article2025년 7월 26일

Ulysses Sequence Parallelism: Training with Million-Token Contexts

Ulysses Sequence Parallelism은 긴 시퀀스 학습에서 시퀀스와 어텐션 헤드를 함께 나누고 all to all 통신으로 재배치해, 단일 GPU 메모리 한계를 넘어 수십만~백만 토큰 문맥 학습을 가능하게 하는 방식이다.

huggingface.co
#ai-architecture#agent-memory#agent-routing#retrieval-index
Model ML is helping financial firms rebuild with AI from the ground up
Article2025년 7월 23일

Model ML is helping financial firms rebuild with AI from the ground up

Model ML은 금융 업무에 특화된 AI 에이전트와 애플리케이션으로 리서치, 분석, 발표자료 작성 같은 복잡한 워크플로를 자동화하며 금융사의 운영 구조 자체를 AI 중심으로 재편하고 있다.

openai.com
#service-design#ai-architecture#multimodal#agent-routing
TimeScope: How Long Can Your Video Large Multimodal Model Go?
Article2025년 7월 23일

TimeScope: How Long Can Your Video Large Multimodal Model Go?

TimeScope는 1분부터 8시간까지의 영상에 짧은 동영상 클립을 삽입해 검색·정보 종합·세밀한 시간 지각 능력을 측정하며, 최신 비전 언어 모델의 장시간 영상 이해가 아직 제한적임을 보여주는 오픈소스 벤치마크다.

huggingface.co
#multimodal#llm#semiconductors#vision-language-models
Pioneering an AI clinical copilot with Penda Health
Article2025년 7월 22일

Pioneering an AI clinical copilot with Penda Health

펜다 헬스가 진료 흐름에 통합한 임상 보조 도구 에이아이 컨설트는 의료진의 통제권을 유지하면서 잠재적 오류를 경고해 진단 오류를 16%, 치료 오류를 13% 상대적으로 줄였다.

openai.com
#agent-deployment#agent-routing#llm#semiconductors
이전1…4142434445…5143 / 51다음