신규 출시 모델과 분야별 인기 오픈소스 모델을 매일 자동 수집합니다. 단가는 100만 토큰당 미국 달러입니다.
총 682건
컨텍스트 400K
입력 $0.025
출력 $0.2
종합 지수 13
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
컨텍스트 400K
입력 $0.05
출력 $0.4
종합 지수 13
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
컨텍스트 131K
입력 $0.03
출력 $0.136
종합 지수 11.6
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
컨텍스트 131K
입력 $0.037
출력 $0.17
종합 지수 11.6
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
컨텍스트 131K
입력 $0.024
출력 $0.112
종합 지수 10
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
컨텍스트 131K
입력 $0.018
출력 $0.09
종합 지수 10
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
컨텍스트 200K
입력 $7.5
출력 $37.5
종합 지수 22.8
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
컨텍스트 200K
입력 $15
출력 $75
종합 지수 22.8
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
좋아요 1,175
다운로드 583K
라이선스 cc-by-4.0
컨텍스트 256K
입력 $0.15
출력 $0.45
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
컨텍스트 256K
입력 $0.3
출력 $0.9
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
컨텍스트 262K
입력 $0.07
출력 $0.28
종합 지수 9.6
Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...
컨텍스트 262K
입력 $0.048
출력 $0.193
종합 지수 7.5
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
좋아요 5
다운로드 58
라이선스 other
컨텍스트 131K
입력 $0.6
출력 $2.2
종합 지수 12.8
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
컨텍스트 131K
입력 $0.13
출력 $0.85
종합 지수 11.1
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
컨텍스트 131K
입력 $0.23
출력 $2.3
종합 지수 12.7
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
컨텍스트 262K
입력 $0.3
출력 $1
종합 지수 11.9
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
컨텍스트 128K
입력 $0.1
출력 $0.2
UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement...
컨텍스트 1M
입력 $0.05
출력 $0.2
종합 지수 10.4
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...