Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#agent-memory
Tag543건YouTube 19Article 524

#agent-memory

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#retrieval-index공동문서 248 · 연관도 66%#context-compression공동문서 286 · 연관도 60%#semiconductors공동문서 457 · 연관도 56%#applications공동문서 432 · 연관도 56%#llm공동문서 468 · 연관도 55%#agent-routing공동문서 201 · 연관도 37%#ai-architecture공동문서 152 · 연관도 30%#agent-deployment공동문서 98 · 연관도 23%#vision-language-models공동문서 64 · 연관도 20%#multimodal공동문서 66 · 연관도 20%
Stargate Community
Article2026년 1월 20일

Stargate Community

OpenAI는 대규모 AI 인프라 확충을 지역사회와의 장기적 협력으로 추진하며, 각 Stargate 캠퍼스가 필요로 하는 전력 비용과 전력망 보강 부담을 자체적으로 책임지겠다고 밝혔다.

openai.com
#service-design#ai-safety#llm#semiconductors
Introducing ChatGPT Go, now available worldwide
Article2026년 1월 16일

Introducing ChatGPT Go, now available worldwide

OpenAI는 저가형 구독제인 ChatGPT Go를 ChatGPT가 제공되는 전 지역으로 확대하며, 무료·Go·Plus·Pro로 이어지는 이용자별 선택 구조를 강화했다.

openai.com
#openai#privacy-design#agent-memory#agent-routing
Dynamic surface codes open new avenues for quantum error correction
Article2026년 1월 13일

Dynamic surface codes open new avenues for quantum error correction

구글 퀀텀 AI는 Willow 프로세서에서 동적 표면 코드를 실험해 결합기 수를 줄이고, 누설로 인한 상관 오류를 억제하며, iSWAP 게이트 기반 오류 정정 가능성을 확인했다.

research.google
#privacy-design#ai-architecture#llm#semiconductors
Hard-braking events as indicators of road segment crash risk
Article2026년 1월 13일

Hard-braking events as indicators of road segment crash risk

구글 리서치는 안드로이드 오토에서 집계한 급제동 사건이 실제 도로 구간 사고율과 유의미한 양의 관계를 보이며, 희소한 사고 기록을 보완하는 도로 안전 선행 지표가 될 수 있음을 검증했다.

research.google
#privacy-design#llm#semiconductors#applications
Next generation medical image interpretation with MedGemma 1.5 and medical speech to text with MedASR
Article2026년 1월 13일

Next generation medical image interpretation with MedGemma 1.5 and medical speech to text with MedASR

Google Research는 의료 영상 해석을 강화한 공개 모델 MedGemma 1.5 4B와 의료 음성 인식 모델 MedASR을 발표하며, 개발자가 의료 AI 애플리케이션을 평가·조정·확장할 수 있는 기반을 넓혔다.

research.google
#privacy-design#multimodal#agent-routing#llm
Zenken boosts a lean sales team with ChatGPT Enterprise
Article2026년 1월 13일

Zenken boosts a lean sales team with ChatGPT Enterprise

일본 기업 젠켄은 챗GPT 엔터프라이즈를 전사적으로 도입해 영업 준비, 전략 분석, 번역, 문서 작성 업무를 줄이고 작은 팀으로도 매출 성장과 생산성 향상을 뒷받침하고 있다.

openai.com
#agent-routing#llm#semiconductors#applications
NeuralGCM harnesses AI to better simulate long-range global precipitation
Article2026년 1월 12일

NeuralGCM harnesses AI to better simulate long-range global precipitation

NeuralGCM은 물리 기반 대기 모델과 NASA 위성 강수 관측으로 학습한 신경망을 결합해 전 지구 강수, 특히 극한 강수와 하루 주기 강수의 장기 시뮬레이션 정확도를 높인 하이브리드 모델이다.

research.google
#llm#semiconductors#applications#agent-deployment
Cohere on Hugging Face Inference Providers 🔥
Article2026년 1월 9일

Cohere on Hugging Face Inference Providers 🔥

코히어가 허깅페이스 허브의 추론 제공자로 합류하면서 기업용 언어·검색·다국어·멀티모달 모델 9종을 웹 화면과 여러 클라이언트 라이브러리에서 서버리스 방식으로 사용할 수 있게 됐다.

huggingface.co
#multimodal#capex-cycle#llm#semiconductors
Introducing ChatGPT Health
Article2026년 1월 7일

Introducing ChatGPT Health

챗지피티 헬스는 의료 기록과 웰니스 앱을 안전하게 연결해 개인화된 건강 정보 이해를 돕되, 진단이나 치료가 아닌 의료진과의 상담을 보조하도록 설계된 전용 공간이다.

openai.com
#privacy-design#ai-architecture#llm#semiconductors
Aligning to What? Rethinking Agent Generalization in MiniMax M2
Article2025년 12월 23일

Aligning to What? Rethinking Agent Generalization in MiniMax M2

MiniMax M2는 벤치마크 성능과 현실적 유용성을 함께 추구하면서, 도구 종류의 확대보다 전체 에이전트 실행 과정에서 발생하는 변화와 교란에 적응하는 능력을 일반화의 핵심으로 삼았다.

huggingface.co
#agent-routing#context-compression#m2-m2#prompt-library
Tokenization in Transformers v5: Simpler, Clearer, and More Modular
Article2025년 12월 18일

Tokenization in Transformers v5: Simpler, Clearer, and More Modular

트랜스포머스 버전 5는 토크나이저의 구조와 학습된 어휘를 분리하고 내부 구성과 클래스 계층을 명확히 해, 토큰화를 더 쉽게 분석·변경·학습할 수 있도록 재설계한다.

huggingface.co
#ai-architecture#agent-routing#workflow-automation#llm
Updating our Model Spec with teen protections
Article2025년 12월 18일

Updating our Model Spec with teen protections

OpenAI는 청소년 이용자를 위해 Model Spec에 U18 원칙을 추가하고, ChatGPT가 13~17세에게 더 안전하고 연령에 맞는 방식으로 응답하도록 보호 장치와 전문가 기반 개선 계획을 강화한다고 밝혔다.

openai.com
#ai-safety#llm#semiconductors#applications
Measuring AI’s capability to accelerate biological research in the wet lab
Article2025년 12월 16일

Measuring AI’s capability to accelerate biological research in the wet lab

OpenAI와 Red Queen Bio는 GPT 5가 실제 습식 실험 반복 과정에서 분자 클로닝 프로토콜을 제안·분석·개선하도록 평가했고, 특정 모델 시스템에서 검증된 클론 회수 효율을 기준 대비 79배 높였다고 보고했다.

openai.com
#agent-routing#context-compression#prompt-library#llm
Gemini-backed Paper Assistant Tool provides automated feedback for theoretical computer scientists at STOC 2026
Article2025년 12월 15일

Gemini-backed Paper Assistant Tool provides automated feedback for theoretical computer scientists at STOC 2026

Google Research는 STOC 2026 제출 논문을 대상으로 Gemini 기반 Paper Assistant Tool을 실험해 이론 컴퓨터과학 논문의 증명 오류, 계산 실수, 논리적 빈틈을 사전 점검하는 가능성을 확인했다.

research.google
#privacy-design#agent-routing#llm#semiconductors
Spotlight on innovation: Google-sponsored Data Science for Health Ideathon across Africa
Article2025년 12월 12일

Spotlight on innovation: Google-sponsored Data Science for Health Ideathon across Africa

Google은 아프리카 전역의 데이터 과학·머신러닝 커뮤니티와 함께 공개 Health AI 모델을 활용한 보건 Ideathon을 열고, 자궁경부암 선별·모성 건강·피부질환 분류 등 지역 의료 과제 해결을 시도한 팀들을 선정·지원했다.

research.google
#ai-distribution#search-advertising#zero-click-search#llm
New in llama.cpp: Model Management
Article2025년 12월 11일

New in llama.cpp: Model Management

llama.cpp 서버의 라우터 모드는 재시작 없이 여러 GGUF 모델을 검색·적재·전환·해제하고, 요청의 모델 지정에 따라 독립 프로세스로 실행되는 모델에 트래픽을 전달합니다.

huggingface.co
#anthropic#ai-architecture#agent-memory#agent-routing
A differentially private framework for gaining insights into AI chatbot use
Article2025년 12월 10일

A differentially private framework for gaining insights into AI chatbot use

Google Research는 AI 챗봇 사용 양상을 분석하면서도 개별 대화가 결과에 과도하게 반영되지 않도록 DP 클러스터링, DP 키워드 추출, LLM 요약을 결합한 Urania 프레임워크를 소개했다.

research.google
#privacy-design#agent-routing#workflow-automation#llm
DeepMath: A lightweight math reasoning Agent with smolagents
Article2025년 12월 8일

DeepMath: A lightweight math reasoning Agent with smolagents

DeepMath는 Qwen3 4B Thinking 기반의 경량 수학 추론 에이전트로, 긴 사고 과정을 짧은 파이썬 실행 단계로 대체하고 GRPO 학습을 통해 더 간결하고 정확한 수학 풀이를 목표로 한다.

huggingface.co
#llm#semiconductors#applications#long-context
Titans + MIRAS: Helping AI have long-term memory
Article2025년 12월 4일

Titans + MIRAS: Helping AI have long-term memory

Titans와 MIRAS는 실행 중 장기 기억을 선택적으로 갱신하는 방식으로, 매우 긴 문맥을 더 빠르고 정확하게 다루려는 새로운 시퀀스 모델링 접근이다.

research.google
#ai-architecture#agent-memory#context-compression#retrieval-index
From Waveforms to Wisdom: The New Benchmark for Auditory Intelligence
Article2025년 12월 3일

From Waveforms to Wisdom: The New Benchmark for Auditory Intelligence

Google Research는 기계 청각 지능을 여덟 가지 핵심 능력으로 표준 평가하는 오픈소스 벤치마크 MSEB를 공개하며, 현재 사운드 임베딩 모델들이 범용성과 견고성에서 큰 성능 여지를 남기고 있다고 밝혔다.

research.google
#multimodal#llm#semiconductors#vision-language-models
Microsoft and Hugging Face expand collaboration
Article2025년 11월 24일

Microsoft and Hugging Face expand collaboration

마이크로소프트와 허깅페이스는 1만 개 이상의 검증된 공개 모델을 애저 AI 파운드리에서 간편하고 안전하게 배포할 수 있도록 협력을 확대했습니다.

huggingface.co
#agent-deployment#llm#semiconductors#applications
OVHcloud on Hugging Face Inference Providers 🔥
Article2025년 11월 24일

OVHcloud on Hugging Face Inference Providers 🔥

OVHcloud가 Hugging Face Inference Provider로 추가되어 사용자는 Hub와 Python·JavaScript SDK에서 다양한 오픈 웨이트 모델을 OVHcloud의 서버리스 추론 환경으로 호출할 수 있게 됐다.

huggingface.co
#service-design#multimodal#agent-routing#llm
AI, networks and Mechanical Turks — Benedict Evans
Article2025년 11월 23일

AI, networks and Mechanical Turks — Benedict Evans

대형 소비자 인터넷 서비스는 사용자 행동의 상관관계로 추천과 발견을 만들어 왔지만, 대규모 언어 모델은 사물과 의도를 더 넓게 해석해 인터넷의 ‘발견 필터’ 자체를 바꿀 수 있다.

Benedict Evans
#llm#semiconductors#applications#agent-memory
One Year Since the “DeepSeek Moment”
Article2025년 11월 21일

One Year Since the “DeepSeek Moment”

딥시크 R1은 기술·도입·심리적 장벽을 낮춰 중국의 개방형 인공지능 생태계를 확장하고, 세계 경쟁의 중심을 개별 모델 성능에서 시스템·응용·생태계 역량으로 이동시킨 전환점이었다.

huggingface.co
#privacy-design#llm#semiconductors#applications
이전1…1314151617…2315 / 23다음