Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#prompt-library
Tag206건YouTube 29Article 177

#prompt-library

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#context-compression공동문서 155 · 연관도 53%#semiconductors공동문서 169 · 연관도 34%#applications공동문서 150 · 연관도 31%#llm공동문서 163 · 연관도 31%#agent-routing공동문서 71 · 연관도 21%#ai-architecture공동문서 48 · 연관도 15%#agent-deployment공동문서 38 · 연관도 14%#privacy-design공동문서 40 · 연관도 13%#agent-memory공동문서 42 · 연관도 12%#service-design공동문서 33 · 연관도 12%
Accelerating Qwen3-8B Agent on Intel® Core™ Ultra with Depth-Pruned Draft Models
Article2025년 4월 30일

Accelerating Qwen3-8B Agent on Intel® Core™ Ultra with Depth-Pruned Draft Models

이 글은 Qwen3 8B를 Intel® Core™ Ultra에서 더 빠르게 실행하기 위해 OpenVINO.GenAI의 추측 디코딩과 깊이 가지치기된 Qwen3 0.6B 드래프트 모델을 결합해 약 1.4배 속도 향상을 얻은 과정을 설명한다.

huggingface.co
#service-design#agent-routing#capex-cycle#context-compression
How to Build an MCP Server with Gradio
Article2025년 4월 30일

How to Build an MCP Server with Gradio

그라디오는 기존 파이썬 함수를 도구로 자동 변환하고 실행 옵션 하나로 웹 인터페이스와 모델 콘텍스트 프로토콜 서버를 함께 제공한다.

huggingface.co
#ai-architecture#agent-deployment#context-compression#prompt-library
Announcing FIRE-1, Our Web Action Agent: Launch Week III - Day 2
Article2025년 4월 15일

Announcing FIRE-1, Our Web Action Agent: Launch Week III - Day 2

Firecrawl은 Launch Week III 2일차에 복잡한 웹사이트를 탐색하고 버튼·검색폼 등과 상호작용해 숨은 데이터를 추출하는 웹 액션 에이전트 FIRE 1을 발표했다.

Eric Ciarla
#agent-routing#context-compression#prompt-library#llm
SmolVLM2: Bringing Video Understanding to Every Device
Article2025년 4월 8일

SmolVLM2: Bringing Video Understanding to Every Device

SmolVLM2는 2.2B·500M·256M 세 가지 크기로 영상 이해의 높은 메모리 효율과 온디바이스 실행 가능성을 제시하며, Transformers와 MLX를 통해 출시 직후부터 다양한 환경에서 활용할 수 있도록 공개된 소형 비전·영상 언어 모델군이다.

huggingface.co
#multimodal#capex-cycle#context-compression#prompt-library
Sharing the latest Model Spec
Article2025년 2월 12일

Sharing the latest Model Spec

OpenAI는 모델 행동 원칙을 정의하는 최신 Model Spec을 공개하며, 사용자·개발자 맞춤화와 지적 자유를 확대하되 실제 위해를 막기 위한 경계와 평가·공개 협업 체계를 함께 강화했다.

openai.com
#openai#privacy-design#agent-deployment#prompt-library
Operator System Card
Article2025년 1월 23일

Operator System Card

오퍼레이터는 화면을 보고 브라우저를 조작하는 컴퓨터 사용 에이전트로, 오픈AI는 실제 행동에서 발생할 수 있는 피해를 줄이기 위해 위험 작업 거부, 중요 행동 전 사용자 확인, 외부 레드팀, 프런티어 위험 평가를 결합한 다층 안전 체계를 적용했다.

openai.com
#openai#privacy-design#multimodal#context-compression
Tiny Agents in Python: a MCP-powered agent in ~70 lines of code
Article2025년 1월 12일

Tiny Agents in Python: a MCP-powered agent in ~70 lines of code

이 글은 Hugging Face의 huggingface hub에 포함된 MCP 클라이언트를 활용해 Python에서 약 70줄짜리 Tiny Agent를 실행하고 구성하며, LLM이 MCP 서버의 도구를 발견·호출·반영하는 흐름을 설명한다.

huggingface.co
#context-compression#prompt-library#llm#semiconductors
Controlling Language Model Generation with NVIDIA's LogitsProcessorZoo
Article2024년 12월 23일

Controlling Language Model Generation with NVIDIA's LogitsProcessorZoo

이 글은 언어 모델의 다음 토큰 선택 과정에서 로짓을 직접 조정해 길이, 문맥 유지, 필수 문구, 선택형 답변 같은 생성 제약을 제어하는 방법을 NVIDIA의 LogitsProcessorZoo와 Hugging Face 생성 API 예시로 설명한다.

huggingface.co
#nvidia#privacy-design#agent-routing#capex-cycle
Welcome to the Falcon 3 Family of Open Models!
Article2024년 12월 17일

Welcome to the Falcon 3 Family of Open Models!

팰컨3는 10억~100억 매개변수 규모에서 과학·수학·코딩·추론 성능과 학습 효율을 함께 높이고, 다양한 배포 형식과 개방형 라이선스를 제공하는 디코더 전용 언어 모델 제품군이다.

huggingface.co
#ai-architecture#llm#semiconductors#applications
How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs
Article2024년 12월 5일

How good are LLMs at fixing their mistakes? A chatbot arena experiment with Keras and TPUs

간단한 일정 관리 대화 실험에서 대규모 언어 모델의 실수 수정 능력을 비교한 결과, 최신 3B~9B급 모델은 대체로 피드백을 반영했지만 소형·구형 모델은 형식 준수와 문맥 유지에서 반복적으로 실패했으며 모델 크기만으로 성능이 보장되지는 않았다.

huggingface.co
#agent-memory#context-compression#retrieval-index#llm
Letting Large Models Debate: The First Multilingual LLM Debate Competition
Article2024년 11월 13일

Letting Large Models Debate: The First Multilingual LLM Debate Competition

BAAI는 정적 벤치마크와 사용자 투표형 아레나의 한계를 보완하기 위해 대형 언어모델들이 다국어로 직접 토론하며 추론력·논증력·언어 능력을 드러내는 FlagEval Debate를 제안한다.

huggingface.co
#anthropic#multimodal#context-compression#prompt-library
Introducing SyGra Studio
Article2024년 4월 4일

Introducing SyGra Studio

SyGra Studio는 합성 데이터 생성 흐름을 시각적으로 구성하고, 생성되는 설정과 실행 상태·비용·결과를 한 화면에서 확인할 수 있게 만든 대화형 작업 환경이다.

huggingface.co
#ai-architecture#agent-routing#context-compression#prompt-library
Explosive growth from AI: A review of the arguments
Article2023년 9월 23일

Explosive growth from AI: A review of the arguments

Epoch AI 글은 고도화된 AI가 인간 노동을 대체할 때 폭발적 경제성장이 가능하다는 성장모형의 논거와, 규제·개발 난도·정렬 문제 같은 반론을 함께 검토하며 가능성은 배제하기 어렵지만 확실하지도 않다고 정리한다.

Ege Erdil
#llm#semiconductors#applications#agent-deployment
AI Canon
Article2023년 5월 25일

AI Canon

a16z의 ‘AI Canon’은 현대 AI를 이해하기 위해 영향력이 컸던 논문, 글, 강의, 실무 가이드, 시장 분석 자료를 단계별로 묶은 선별 참고 목록이다.

Derrick Harris, Matt Bornstein, Guido Appenzeller
#anthropic#privacy-design#ai-architecture#agent-routing
이전1…7899 / 9다음