신규 출시 모델과 분야별 인기 오픈소스 모델을 매일 자동 수집합니다. 단가는 100만 토큰당 미국 달러입니다.
총 682건
컨텍스트 1M
입력 $10
출력 $50
종합 지수 49.6
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
컨텍스트 128K
입력 무료
출력 무료
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
컨텍스트 131K
입력 $0.2
출력 $0.2
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
좋아요 791
다운로드 104K
라이선스 other
컨텍스트 1M
입력 무료
출력 무료
종합 지수 22.9
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
컨텍스트 262K
입력 $0.5
출력 $2.2
종합 지수 22.9
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
컨텍스트 1M
입력 $0.32
출력 $1.28
종합 지수 25.2
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
컨텍스트 1M
입력 $0.3
출력 $1.2
종합 지수 29.2
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
좋아요 71
다운로드 24K
라이선스 other
컨텍스트 262K
입력 $0.2
출력 $1.15
종합 지수 19.5
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
컨텍스트 1M
입력 $2.5
출력 $12.5
종합 지수 41.8
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
컨텍스트 1M
입력 $5
출력 $25
종합 지수 41.8
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
컨텍스트 1M
입력 $1.48
출력 $4.42
종합 지수 29.5
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
컨텍스트 256K
입력 $1
출력 $2
종합 지수 27.2
Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...
컨텍스트 1M
입력 $0.75
출력 $4.5
종합 지수 33.6
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
컨텍스트 1M
입력 $1.5
출력 $9
종합 지수 33.6
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
좋아요 1,168
다운로드 1.3M
라이선스 other
컨텍스트 33K
입력 $0.15
출력 $1.5
Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...
컨텍스트 1M
입력 $0.125
출력 $0.75
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
컨텍스트 1M
입력 $0.25
출력 $1.5
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...