모델을 불러오고 있습니다…
모델을 불러오고 있습니다…
총 247개 중 121–150개
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...
GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...
MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...
Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.
LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.2](/models/sao10k/l3-euryale-70b).
Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...
This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
업스테이지의 프런티어급 한국어 모델. 문서 이해·RAG 파이프라인에서 높은 정확도.
MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly 10x...
Anthropic 최상위 모델. 긴 문서 분석·정교한 글쓰기·코드 리뷰에 강합니다.
구글 제미나이 플래시. 속도와 비용의 균형이 좋은 범용 모델. 100만 토큰 컨텍스트.
구글 제미나이 프로. 구글의 상위 추론 모델. 100만 토큰 컨텍스트.
구글 제미나이 플래시 라이트. 가장 싼 구글 모델. 대량 분류·요약에 맞다. 100만 토큰 컨텍스트. 프리뷰 — 예고 없이 바뀔 수 있다.
구글 제미나이 프로. 구글의 상위 추론 모델. 100만 토큰 컨텍스트. 프리뷰 — 예고 없이 바뀔 수 있다. 커스텀 도구 호출 변형.
구글 제미나이 플래시. 속도와 비용의 균형이 좋은 범용 모델. 100만 토큰 컨텍스트.
구글 제미나이 플래시 라이트. 가장 싼 구글 모델. 대량 분류·요약에 맞다. 100만 토큰 컨텍스트.
구글 제미나이 플래시. 속도와 비용의 균형이 좋은 범용 모델. 100만 토큰 컨텍스트.
구글 제미나이 플래시. 속도와 비용의 균형이 좋은 범용 모델. 100만 토큰 컨텍스트.
초저지연 음성 합성. 약 75ms 로 응답해 실시간 대화에 맞다. 32개 언어. 한 번에 40,000자까지.
속도와 품질의 균형. 플래시보다 표현이 낫고 다국어를 지원한다. 32개 언어. 한 번에 40,000자까지.
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents. It accepts text, image, video, and audio input, and answers with text plus optional structured annotations: points, boxes, polygons, tracks,...
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...