ローカルLLM のリリース年表
主な「ローカルLLM」のリリース年表をまとめました。
2022年
11月30日 OpenAI - ChatGPT リリース
2023年
2月24日 LLaMA
7B、13B、33B、65B を研究者向けの限定リリース。
3月13日 Alpaca
Instructionデータセットによる学習効果が注目される。
3月14日 OpenAI - GPT-4 リリース
3月19日 Llama.cpp
Llamaの高速推論が注目される。
3月20日 Alpaca-LoRA
LoRAでのInstructionデータセットによる学習効果が注目される。
3月30日 Vicuna
ShareGPT (ChatGPTログ) による学習効果が注目される。
4月17日 RedPajama-Data-1T
オープンなLlamaを実装するため大規模データセットが作成される。
5月15日 RWKV-Raven
RWKV
・RWKV/rwkv-raven-14b
・RWKV/rwkv-raven-7b
・RWKV/rwkv-raven-3b
・RWKV/rwkv-raven-1b5
・RWKV/rwkv-4-14b-pile
・RWKV/rwkv-4-7b-pile
・RWKV/rwkv-4-3b-pile
・RWKV/rwkv-4-1b5-pile
・RWKV/rwkv-4-430m-pile
・RWKV/rwkv-4-169m-pile
5月17日 Rinna-3.6B
・rinna/japanese-gpt-neox-3.6b
・rinna/japanese-gpt-neox-3.6b-instruction-sft
5月18日 OpenCALM
・cyberagent/open-calm-7b
・cyberagent/open-calm-3b
・cyberagent/open-calm-1b
・cyberagent/open-calm-large
・cyberagent/open-calm-small
・cyberagent/open-calm-medium
5月31日 Rinna-3.6B-instruction-ppo
・rinna/japanese-gpt-neox-3.6b-instruction-ppo
7月14日 RWKV-4-World
RWKV
・BlinkDL/rwkv-4-world
7月17日 OpenAI - Code Interpreter リリース
7月19日 LLaMA 2
・meta-llama/Llama-2-7b-hf
・meta-llama/Llama-2-13b-hf
・meta-llama/Llama-2-70b-hf
・meta-llama/Llama-2-7b-chat-hf
・meta-llama/Llama-2-13b-chat-hf
・meta-llama/Llama-2-70b-chat-hf
・meta-llama/Llama-2-7b
・meta-llama/Llama-2-13b
・meta-llama/Llama-2-70b
・meta-llama/Llama-2-7b-chat
・meta-llama/Llama-2-13b-chat
・meta-llama/Llama-2-70b-chat
7月31日 Rinna-4B
・rinna/bilingual-gpt-neox-4b
・rinna/bilingual-gpt-neox-4b-8k
・rinna/bilingual-gpt-neox-4b-instruction-sft
・rinna/bilingual-gpt-neox-4b-instruction-ppo
・rinna/bilingual-gpt-neox-4b-minigpt4
8月10日 Japanese StableLM Alpha-7B
・stabilityai/japanese-stablelm-base-alpha-7b
・stabilityai/japanese-stablelm-instruct-alpha-7b
8月11日 AIBunCho-6B
・AIBunCho/japanese-novel-gpt-j-6b
8月14日 Line-3.6B
・line-corporation/japanese-large-lm-3.6b
8月17日 Japanese InstructBLIP Alpha
・stabilityai/japanese-instructblip-alpha
8月18日 Line-3.6B-instruction-sft
・line-corporation/japanese-large-lm-3.6b-instruction-sft
8月22日 WebLab-10B
・matsuo-lab/weblab-10b
・matsuo-lab/weblab-10b-instruction-sft
8月24日 CodeLlama
Code
・codellama/CodeLlama-34b-hf
・codellama/CodeLlama-34b-Instruct-hf
・codellama/CodeLlama-34b-Python-hf
・codellama/CodeLlama-13b-hf
・codellama/CodeLlama-13b-Instruct-hf
・codellama/CodeLlama-13b-Python-hf
・codellama/CodeLlama-7b-hf
・codellama/CodeLlama-7b-Instruct-hf
・codellama/CodeLlama-7b-Python-hf
8月29日 ELYZA-7B
・elyza/ELYZA-japanese-Llama-2-7b-instruct
・elyza/ELYZA-japanese-Llama-2-7b-fast-instruct
・elyza/ELYZA-japanese-Llama-2-7b
・elyza/ELYZA-japanese-Llama-2-7b-fast
9月6日 Open Interpreter
ローカル環境で、OpenAIの「Code Interpreter」相当の機能を実行可能になる。
9月6日 Falcon 180B
ローカルLLM初の180BでGPT-3 (175B) 以上のパラメータになる。
・tiiuae/falcon-180B
9月7日 Heron
VLM
・turing-motors/heron-preliminary-git-Llama-2-70b-v0
・turing-motors/heron-chat-blip-ja-stablelm-base-7b-v0
・turing-motors/heron-chat-git-ELYZA-fast-7b-v0
・turing-motors/heron-chat-git-ja-stablelm-base-7b-v0
9月21日 Xwin-LM
ベンチマーク「AlpacaEval」で「GPT-4」を追い抜き1位を獲得したモデル。
・Xwin-LM/Xwin-LM-70B-V0.1
・Xwin-LM/Xwin-LM-13B-V0.1
・Xwin-LM/Xwin-LM-7B-V0.1
9月25日 OpenAI - GPT-4Vリリース
9月27日 Mistral-7B-v0.1
「Mistral AI」が開発したLLM。7Bにもかわらず「Llama 2 13B」や「Llama 1 34B」など大きなモデルよりもベンチマークで高スコアを獲得したモデル。
・mistralai/Mistral-7B-v0.1
・mistralai/Mistral-7B-Instruct-v0.1
9月28日 PLaMo-13B
・pfnet/plamo-13b
10月3日 Qwen-14B
・Qwen/Qwen-14B
・Qwen/Qwen-14B-Chat
10月5日 LLaVA-1.5
VLM
・liuhaotian/llava-v1.5-13b
10月10日 Japanese StableLM Instruct Alpha-7B-v2
・stabilityai/japanese-stablelm-instruct-alpha-7b-v2
10月20日 LLM-jp-13B
・llm-jp-13b-instruct-full-jaster-v1.0
・llm-jp-13b-instruct-full-jaster-dolly-oasst-v1.0
・llm-jp-13b-instruct-full-dolly-oasst-v1.0
・llm-jp-13b-instruct-lora-jaster-v1.0
・llm-jp-13b-instruct-lora-jaster-dolly-oasst-v1.0
・llm-jp-13b-instruct-lora-dolly-oasst-v1.0
・llm-jp-13b-v1.0
・llm-jp-1.3b-v1.0
10月25日 Japanese Stable LM 3B-4E1T
・Japanese Stable LM 3B-4E1T Base
・Japanese Stable LM 3B-4E1T Instruct
10月25日 Japanese Stable LM Gamma 7B
・Japanese Stable LM Base Gamma 7B
・Japanese Stable LM Instruct Gamma 7B
10月26日 Stockmark-13B
・stockmark/stockmark-13b
10月27日 Zephyr-7B-Beta
・HuggingFaceH4/zephyr-7b-beta
10月25日 RWKV-5-World
RWKV
10月31日 Youri-7B
・rinna/youri-7b
・rinna/youri-7b-instruction
・rinna/youri-7b-chat
・rinna/youri-7b-gptq
・rinna/youri-7b-instruction-gptq
・rinna/youri-7b-chat-gptq
11月2日 Japanese Stable LM Beta
・stabilityai/japanese-stablelm-base-beta-7b
・stabilityai/japanese-stablelm-base-beta-70b
・stabilityai/japanese-stablelm-instruct-beta-7b
・stabilityai/japanese-stablelm-instruct-beta-70b
・stabilityai/japanese-stablelm-base-ja_vocab-beta-7b
・stabilityai/japanese-stablelm-instruct-ja_vocab-beta-7b
11月2日 CALM2
・cyberagent/calm2-7b
・cyberagent/calm2-7b-chat
11月6日 OpenAI DevDay
11月4日 DeepSeek Coder
Code
・deepseek-ai/deepseek-coder-33b-instruct
・deepseek-ai/deepseek-coder-33b-base
・deepseek-ai/deepseek-coder-6.7b-instruct
・deepseek-ai/deepseek-coder-6.7b-base
・deepseek-ai/deepseek-coder-5.7bmqa-base
・deepseek-ai/deepseek-coder-1.3b-instruct
・deepseek-ai/deepseek-coder-1.3b-base
11月7日 PLaMo-13B-Instruct
・pfnet/plamo-13b-instruct
・pfnet/plamo-13b-instruct-nc
11月13日 Japanese Stable VLM
VLM
11月15日 ELYZA-japanese-CodeLlama-7b
Code
・elyza/ELYZA-japanese-CodeLlama-7b
・elyza/ELYZA-japanese-CodeLlama-7b-instruct
11月15日 Japanese Stable CLIP
・stabilityai/japanese-stable-clip-vit-l-16
11月28日 Starling-7B
・berkeley-nest/Starling-LM-7B-alpha
11月30日 DeepSeek LLM
・deepseek-ai/deepseek-llm-67b-chat
・deepseek-ai/deepseek-llm-67b-base
・deepseek-ai/deepseek-llm-7b-chat
・deepseek-ai/deepseek-llm-7b-base
12月1日 Qwen-72B・Qwen-Audio
・Qwen/Qwen-72B
・Qwen/Qwen-72B-Chat
・Qwen/Qwen-Audio-Chat
12月6日 Shisa-7B
・augmxnt/shisa-base-7b-v1
・augmxnt/shisa-7b-v1
12月8日 StableLM Zephyr 3B
・stabilityai/stablelm-zephyr-3b
12月8日 StripedHyena-7B
Hyena
・togethercomputer/StripedHyena-Hessian-7B
・togethercomputer/StripedHyena-Nous-7B
12月9日 Mixtral-8x7b-v0.1
・mistralai/Mixtral-8x7B-v0.1
・mistralai/Mixtral-8x7B-Instruct-v0.1
12月11日 Mistral-7B-Instruct-v0.2
・mistralai/Mistral-7B-Instruct-v0.2
※ Mistral-7B-v0.1ベースのInstructモデルのv0.2
12月13日 phi-2
・microsoft/phi-2
12月19日 Swallow
・tokyotech-llm/Swallow-7b-hf
・tokyotech-llm/Swallow-7b-instruct-hf
・tokyotech-llm/Swallow-13b-hf
・tokyotech-llm/Swallow-13b-instruct-hf
・tokyotech-llm/Swallow-70b-hf
・tokyotech-llm/Swallow-70b-instruct-hf
12月19日 PowerInfer
Inference
12月21日 Nekomata
・rinna/nekomata-14b
・rinna/nekomata-14b-instruction
・rinna/nekomata-7b
・rinna/nekomata-7b-instruction
12月27日 ELYZA-japanese-Llama-2-13B
・elyza/ELYZA-japanese-Llama-2-13b
・elyza/ELYZA-japanese-Llama-2-13b-instruct
・elyza/ELYZA-japanese-Llama-2-13b-fast
・elyza/ELYZA-japanese-Llama-2-13b-fast-instruct
12月29日 Karasu・Qarasu
・lightblue/qarasu-14B-chat-plus-unleashed
・lightblue/karasu-7B-chat-plus-unleashed
・lightblue/karasu-7B-chat
・lightblue/karasu-7B
2024年
1月3日 M2UGen
MLM
1月10日 Phixtral
・mlabonne/phixtral-4x2_8
・mlabonne/phixtral-2x2_8
🔀 Phixtral
— Maxime Labonne (@maximelabonne) January 9, 2024
I made the first efficient Mixture of Experts with phi-2 models. 🥳
It combines 2 to 4 fine-tuned models and is better than each individual expert.
🤗 phixtral-2x2_8: https://t.co/XbPpsF76vN
🤗 phixtral-4x2_8: https://t.co/9xfRd46585 pic.twitter.com/coRpRIxG2V
1月16日 Stable Code 3B
Code
1月20日 StableLM 2 1.6B
1月22日 Stable LM 2 1.6B
1月23日 Orion-14B
・OrionStarAI/Orion-14B-Base
・OrionStarAI/Orion-14B-Chat
・OrionStarAI/Orion-14B-LongChat
・OrionStarAI/Orion-14B-Chat-RAG
・OrionStarAI/Orion-14B-Chat-Plugin
・OrionStarAI/Orion-14B-Base-Int4
・OrionStarAI/Orion-14B-Chat-Int4
1月23日 Yi-VL-34B
1月29日 RWKV-Eagle-7B
RWKV
1月29日 CodeLlama-70B
Code
・codellama/CodeLlama-70b-hf
・codellama/CodeLlama-70b-Instruct-hf
・codellama/CodeLlama-70b-Python-hf
1月30日 LLaVA-1.6
VLM
1月31日 KARAKURI LM
2月4日 Qwen1.5
2月7日 MobileVLM V2
・mtgv/MobileVLM_V2-7B
・mtgv/MobileVLM_V2-3B
・mtgv/MobileVLM_V2-1.7B
2月9日 LLM-jp 13B v1.1
2月19日 kotomamba
Mamba
2月21日 Gemma
2月28日 StarCoder 2
Code
2月28日 BitNet
3月3日 Swallow-7B-plus
3月6日 heron-blip-v1
VLM
3月11日 Swallow-MS 7B
3月11日 Swallow-MX 8x7B
3月11日 Command R
3月17日 Grok-1
3月21日 EvoVLM-JP-v1
VLM
3月21日 EvoLLM-JP-v1
3月21日 RakutenAI-7B
3月22日 ao-Karasu-72B
3月24日 Mistral-7B-v0.2
Mistral just announced at @SHACK15sf that they will release a new model today:
— Marvin von Hagen (@marvinvonhagen) March 23, 2024
Mistral 7B v0.2 Base Model
- 32k instead of 8k context window
- Rope Theta = 1e6
- No sliding window pic.twitter.com/iAuEUEOw5K
3月27日 DBRX
3月28日 Qwen1.5-MoE
3月28日 Jamba
Mamba
4月2日 Qwen1.5-32B
4月3日 LightChatAssistant
4月4日 Command R+
4月5日 JetMoE-8B
4月5日 Gemma-1.1
4月8日 Stable LM 2 12B
4月9日 CodeGemma
Code
4月9日 RecurrentGemma
4月15日 Idefics2
VLM
4月15日 Japanese-Starling-ChatV-7B
4月16日 WizardLM-2 8x22B・70B・7B
🔥Today we are announcing WizardLM-2, our next generation state-of-the-art LLM.
— WizardLM (@WizardLM_AI) April 15, 2024
New family includes three cutting-edge models: WizardLM-2 8x22B, 70B, and 7B - demonstrates highly competitive performance compared to leading proprietary LLMs.
📙Release Blog:… pic.twitter.com/bclr4aBib1
4月17日 Mixtral-8x22B-v0.1
4月18日 Llama 3
4月23日 Suzume-Llama-3-8B
4月23日 Phi-3-mini
4月24日 OpenELM
・apple/OpenELM-270M
・apple/OpenELM-450M
・apple/OpenELM-1_1B
・apple/OpenELM-3B
・apple/OpenELM-270M-Instruct
・apple/OpenELM-450M-Instruct
・apple/OpenELM-1_1B-Instruct
・apple/OpenELM-3B-Instruct
4月24日 LEIA
4月24日 Snowflake Arctic
4月24日 Antler-7B-Novel-Writing
・Aratako/Antler-7B-Novel-Writing
・Aratako/Antler-7B-Novel-Writing-GGUF
4月26日 SniffyOtter-7B-Novel-Writing-NSFW
・Aratako/SniffyOtter-7B-Novel-Writing-NSFW
・Aratako/SniffyOtter-7B-Novel-Writing-NSFW-GGUF
4月26日 Qwen/Qwen1.5-110B
4月26日 Swallow-MS-7b-instruct v0.1
4月29日 StarCoder2-Instruct
Code
4月30日 LLM-jp-13B v2.0
・llm-jp/llm-jp-13b-instruct-full-ac_001_16x-dolly-ichikara_004_001_single-oasst-oasst2-v2.0
・llm-jp/llm-jp-13b-instruct-full-ac_001-dolly-ichikara_004_001_single-oasst-oasst2-v2.0
・llm-jp/llm-jp-13b-instruct-full-dolly-ichikara_004_001_single-oasst-oasst2-v2.0
・llm-jp/llm-jp-13b-v2.0
5月1日 Llama-3-Youko-8B
5月1日 Ninja-v1 ・ Vecteus-v1
5月3日 Assistance
5月7日 DeepSeek-V2
5月7日 KARAKURI LM 8x7B Chat v0.1
KARAKURI LM 8x7B Chat v0.1を公開しました!
— Tomofumi Nakayama (@txmy) May 7, 2024
model: https://t.co/bJJ9Tad1mH
demo: https://t.co/QlWZ8W2i9n
AWS Trainiumで学習されたMoEモデルとしては多分世界初です。
詳細はスレッドに。 pic.twitter.com/2wuBGPZJL5
5月7日 KARAKURI LM 7B APM v0.1
5月9日 Japanese Stable LM 2 1.6B
・stabilityai/japanese-stablelm-2-base-1_6b
・stabilityai/japanese-stablelm-2-instruct-1_6b
5月9日 ArrowPro-7B-KUJIRA
5月10日 ArrowPro-7B-RobinHood
5月10日 Ocuteus-v1
5月10日 Fugaku-LLM-13B
・Fugaku-LLM/Fugaku-LLM-13B
・Fugaku-LLM/Fugaku-LLM-13B-instruct
5月13日 OpenAI - GPT-4o リリース
5/13 Yi-1.5
5月14日 PaliGemma
VLM
5月16日 Stockmark-100b
5月20日 MiniCPM-Llama3-V 2.5
VLM
5月21日 Phi-3-small (7B)
5月21日 Phi-3-medium (14B)
・microsoft/Phi-3-medium-128k-instruct
・microsoft/Phi-3-medium-4k-instruct
5月21日 Phi-3-vision
VLM
5月21日 Ninja-v1-RP-expressive
5月22日 Mistral-7B-v0.3
・mistralai/Mistral-7B-Instruct-v0.3
・mistralai/Mistral-7B-v0.3
5月23日 Aya-23
5月26日 ArrowPro-7B-KillerWhale
5月29日 Codestral-22B-v0.1
Code
5月29日 Umievo-itr012-Gleipnir-7B
6月1日 Tanuki-8B
・hatakeyama-llm-team/Tanuki-8B
・hatakeyama-llm-team/Tanuki-8B-Instruct
・hatakeyama-llm-team/Tanuki-8B-Instruct-without-DPO
6月1日 Oumuamua-7B
6月5日 GLM-4-9B
6月7日 Qwen2
6月12日 RecurrentGemma-9B
6月13日 Llava Calm2 Siglip
VLM
6月14日 Nemotron-4-340B
6月14日 Sarashina1
6月14日 Sarashina2
6月20日 KARAKURI LM 8x7B Instruct v0.1
6月26日 Llama-3-ELYZA-JP-8B
6月27日 Gemma 2
7月1日 Llama-3-Swallow
7月3日 CALM3-22B-Chat
7月3日 internLM 2.5
7月3日 InternLM-XComposer2.5
VLM
7月16日 Mathstral-7B-v0.1
Math
7月16日 Mamba-Codestral-7B-v0.1
Mamba、Code
7月18日 Mistral NeMo
7月19日 DCLM-7B
7月19日 Athene-70B
7月23日 Llama-3.1-405B・70B・8B
7月24日 Mistral Large 2
7月25日 Llama 3 Youko
7月26日 Llama-3.1-70B-Japanese-Instruct-2407
7月30日 Llama-3.1-70B-EZO-1.1-it ・ Llama-3.1-8B-EZO-1.1-it
7月31日 Gemma 2 2B
7月31日 ShieldGemma
Moderation
8月1日 EZO-Common-T2-2B-gemma-2-it
8月2日 Llama-3-EvoVLM-JP-v2
VLM
8月5日 Llama-3-EZO-VLM-1
VLM
8月7日 MiniCPM-V2.6
VLM
8月7日 Sarashina2-70B
8月12日 FalconMamba 7B
Mamba
8月13日 LongWriter
8月19日 EZO-InternVL2-26B
VLM
8月20日 Phi-3.5-mini-instruct
8月20日 Phi-3.5-MoE-instruct
8月20日 Phi-3.5-vision-instruct
VLM
8月21日 Borea-Phi-3.5-mini-Instruct
8月22日 Jamba-1.5
Mamba
8月29日 Qwen2-VL
VLM
8月30日 Tanuki-8x8B
8月30日 Command-R-plus-08-2024 ・ Command-R-08-2024
9月6日 DeepSeek-V2.5
・deepseek-ai/DeepSeek-V2.5
9月11日 Reader-LM
HTML-to-Markdown
9月11日 Pixtral-12B
VLM
・mistralai/Pixtral-12B-2409
9月11日 LLaMA-Omni
Speech-to-Speech
・ICTNLP/Llama-3.1-8B-Omni
9月12日 DataGemma
9月17日 LLM-jp-3 172B beta1
9月17日 Mistral-Small-Instruct-2409
9月18日 CogVideoX-5b-I2V
9月18日 Qwen2.5
9月18日 Moshi
9月19日 Kurage
9月24日 EZO-Qwen2.5 ・ EZO-AutoCoTRAG-Qwen2.5
9月25日 LLM-jp-3 1.8B・3.7B・13B
9月25日 Llama 3.2 Vision
VLM
9月25日 Llama 3.2 1B・3B
・meta-llama/Llama-3.2-1B
・meta-llama/Llama-3.2-1B-Instruct
・meta-llama/Llama-3.2-3B
・meta-llama/Llama-3.2-3B-Instruct
9月25日 Molmo
VLM
9月30日 llm-jp-3-3.7b-instruct-EZO-Humanities ・ llm-jp-3-3.7b-instruct-EZO-Common
・AXCXEPT/llm-jp-3-3.7b-instruct-EZO-Humanities
・AXCXEPT/llm-jp-3-3.7b-instruct-EZO-Common
10月3日 Gemma 2 Baku 2B
10月3日 Gemma 2 JPN
10月8日 Llama-3.1-Swallow v0.1
10月10日 Entropix
10月10日 ARIA
Multimodal Native MoE Model
10月15日 Ichigo Llama 3.1
Real Time Voice AI
10月15日 Zamba2-7B-Instruct
Mamba
10月15日 PLaMo-100B
10月16日 Gemma-ASP
10月16日 Ministral
10月16日 Llama-3.1-Nemotron-70B
10月18日 bitnet.cpp
BitNet
10月18日 Janus-1.3B
マルチモーダル理解(Image+Text→Text)と生成(Text→Image)の両対応
10月18日 Meta Spirit LM
10月21日 Granite 3.0
10月24日 Aya Expanse
10月31日 SmolLM2
11月8日 Sarashina2-8x70B
11月11日 Llama-3.1-Swallow v0.2
11月12日 Qwen2.5-Coder 32B
11月14日 Athene-V2 72B
11月15日 LLM-jp-3 172B beta2
11月18日 Pixtral Large ・ Mistral Large 2211
11月20日 LLM-jp-3 VILA 14B
11月28日 QwQ-32B-Preview
12月6日 Llama 3.3 70B
As we continue to explore new post-training techniques, today we're releasing Llama 3.3 — a new open source model that delivers leading performance and quality across text-based use cases such as synthetic data generation at a fraction of the inference cost. pic.twitter.com/BNoV2czGKL
— AI at Meta (@AIatMeta) December 6, 2024
12月6日 Qwen2-VL-72B
12月10日 Sarashina2.1-1B ・ Sarashina-Embedding-v1-1B
・sbintuitions/sarashina2.1-1b
・sbintuitions/sarashina-embedding-v1-1b
12月11日 Sarashina2.1-1B-SFT
12月13日 Phi-4
・Introducing Phi-4: Microsoft’s Newest Small Language Model Specializing in Complex Reasoning
12月13日 DeepSeek-VL2
・deepseek-ai/deepseek-vl2
・deepseek-ai/deepseek-vl2-small
・deepseek-ai/deepseek-vl2-tiny
12月14日 Command-R7B
12月16日 Apollo
12月17日 Falcom 3
12月23日 Llama-3.1-Swallow-8B-Instruct-v0.3
12月24日 llm-jp-3-172b-instruct3
12月25日 QvQ-72B-Preview
12月25日 DeepSeek-V3
12月30日 Llama-3.1-Swallow-70B-Instruct-v0.3
2025年
1月8日 phi-4 (MIT License)
1月14日 MiniMax-Text-01・MiniMax-VL-01
1月20日 DeepSeek-R1
🚀 DeepSeek-R1 is here!
— DeepSeek (@deepseek_ai) January 20, 2025
⚡ Performance on par with OpenAI-o1
📖 Fully open-source model & technical report
🏆 MIT licensed: Distill & commercialize freely!
🌐 Website & API are live now! Try DeepThink at https://t.co/v1TFy7LHNy today!
🐋 1/n pic.twitter.com/7BlpWAPu6y
1月20日 DeepSeek-R1-Zero
1月20日 DeepSeek-R1-Distill
・deepseek-ai/DeepSeek-R1-Distill-Qwen-32B
・deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
・deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
・deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
・deepseek-ai/DeepSeek-R1-Distill-Llama-70B
・deepseek-ai/DeepSeek-R1-Distill-Llama-8B
1月24日 KARAKURI LM 32B Thinking 2501 Experimental
先月実験でちょろっとやってたQwQの日本語モデルです。https://t.co/GKtdMUrJx3
— Tomofumi Nakayama (@txmy) January 23, 2025
一貫した日本語で推論できますが、細かいチューニングはしてないので繰り返しが時々発生します。
使う理由は32Bクラスで一貫した日本語で推論内容を見たいときくらいしかないかと思いますが、DeepSeek…
1月24日 J-Moshi
1月27日 Qwen2.5-1M
1月27日 DeepSeek-R1-Distill-Qwen-14B-Japanese
・cyberagent/DeepSeek-R1-Distill-Qwen-32B-Japanese
・cyberagent/DeepSeek-R1-Distill-Qwen-14B-Japanese
1月27日 ABEJA-Qwen2.5-32b-Japanese-v0.1
1月27日 phi-4-open-R1-Distill-EZOv1
1月27日 DeepSeek-R1-GGUF 1.58bit
1月27日 Janus-Pro
1月27日 Qwen2.5-VL
1月29日 DeepSeek-R1-Distill-Qwen-7B-Japanese
1月30日 TinySwallow
1月31日 Mistral Small 3
2月5日 LLM-jp-3 instruct3 150M・440M・980M・7.2B
2月8日 PLaMo-2-1B
2月12日 RakutenAI-2.0
2月13日 Qwen2.5 Bakeneko 32B
2月19日 R1 1776
2月25日 PLaMo-2-8B
2月25日 Asagi
3月5日 Aya Vision
3月5日 Sarashina2.2
・sbintuitions/sarashina2.2-3b-instruct-v0.1
・sbintuitions/sarashina2.2-1b-instruct-v0.1
・sbintuitions/sarashina2.2-0.5b-instruct-v0.1
・sbintuitions/sarashina2.2-3b
・sbintuitions/sarashina2.2-1b
・sbintuitions/sarashina2.2-0.5b
3月5日 QwQ-32B
3月10日 Llama 3.3 Swallow
3月10日 いざなみ・くしなだ
3月12日 Gemma 3
3月13日 Command A
3月13日 QwQ Bakeneko 32B
3月17日 Sarashina2-Vision
3月17日 Mistral Small 3.1
3月19日 Qwen2.5 Bakeneko 32B Instruct V2
3月24日 DeepSeek-V3-0324
3月24日 Qwen2.5-VL-32B
4月5日 Llama 4 Scout・Llama 4 Maverick
4月16日 Shisa V2
4月17日 ABEJA QwQ-32B Reasoning Model
4月28日 Qwen3
4月30日 Phi-4-reasoning ・ Phi-4-reasoning-plus ・ Phi-4-mini-reasoning
5月1日 ELYZA-Thinking-1.0-Qwen-32B ・ ELYZA-Shortcut-1.0-Qwen-32B/7B
5月12日 Heron-NVILA-Lite
5月19日 Gemma-2-Llama Swallow
5月20日 Gemma 3n
5月21日 Devstral
5月27日 PLaMo翻訳
5月28日 DeepSeek-R1-0528 ・ DeepSeek-R1-0528-Qwen3-8B
5月30日 LLM-jp-3.1 シリーズ instruct4
6月3日 Stockmark-2-VL-100B
6月3日 Llama 3.1 Shisa V2 405B
🚀https://t.co/liBSUWXYls 、国産最高性能の多言語LLM『Llama 3.1 Shisa V2 405B』をオープンソースで公開!
— shisa.ai (@shisa_ai) June 3, 2025
GPT-4/Turboを超える日本語性能、主要ベンチマークでGPT-4oやDeepSeek-V3と同等の結果を達成。
日本語専用ベンチ3種も近日公開!
今すぐチェック: https://t.co/F65yg1amIH#ShisaAI #LLM pic.twitter.com/avmFa3kt0b
6月10日 Magistral-Small-2506
6月10日 LiteRT-LM
Run Gemma 3n on desktop (Mac, Win, Linux) and IoT with LiteRT-LM. This early preview is offered with a C++ API, and is fully open source and customizable. pic.twitter.com/rGYolgT1sy
— Google AI Developers (@googleaidevs) June 10, 2025
6月10日 PLaMo翻訳CLI
6月16日 MiniMax-M1
Day 1/5 of #MiniMaxWeek: We’re open-sourcing MiniMax-M1, our latest LLM — setting new standards in long-context reasoning.
— MiniMax (official) (@MiniMax__AI) June 16, 2025
- World’s longest context window: 1M-token input, 80k-token output
- State-of-the-art agentic use among open-source models
- RL at unmatched efficiency:… pic.twitter.com/bGfDlZA54n
6月25日 Llama 3.1 Swallow 8B v0.5
7月8日 SmolLM3
7月9日 KARAKURI VL
7月11日 Kimi K2
7月21日 Qwen3-235B-A22B-Instruct-2507
7月22日 Qwen3-Coder-480B-A35B-Instruct
7月25日 Qwen3-235B-A22B-Thinking-2507
7月29日 Qwen3-30B-A3B
7月30日 Qwen3-30B-A3B-Thinking-2507
7月31日 Qwen3-Coder-30B-A3B-Instruct
8月5日 gpt-oss-120b・gpt-oss-20b
8月6日 Qwen3-4B-Instruct-2507 ・ Qwen3-4B-Thinking-2507
8月19日 DeepSeek-V3.1-Base
8月20日 Sarashina-embedding-v2-1B
8月21日 DeepSeek-V3.1
8月23日 Grok-2.5 - オープンウェイト
9月8日 qep-1bit-extreme
9月12日 Qwen3-Next-80B-A3B
9月17日 Granite-Docling-258M
9月22日 DeepSeek-V3.1-Terminus
🚀 DeepSeek-V3.1 → DeepSeek-V3.1-Terminus
— DeepSeek (@deepseek_ai) September 22, 2025
The latest update builds on V3.1’s strengths while addressing key user feedback.
✨ What’s improved?
🌐 Language consistency: fewer CN/EN mix-ups & no more random chars.
🤖 Agent upgrades: stronger Code Agent & Search Agent performance.…
9月22日 Qwen3-Omni
🚀 Introducing Qwen3-Omni — the first natively end-to-end omni-modal AI unifying text, image, audio & video in one model — no modality trade-offs!
— Qwen (@Alibaba_Qwen) September 22, 2025
🏆 SOTA on 22/36 audio & AV benchmarks
🌍 119L text / 19L speech in / 10L speech out
⚡ 211ms latency | 🎧 30-min audio… pic.twitter.com/qGn34N7Xvd
9月23日 Qwen3-VL
9月23日 Qwen3-Max
9月29日 DeepSeek-V3.2-Exp
🚀 Introducing DeepSeek-V3.2-Exp — our latest experimental model!
— DeepSeek (@deepseek_ai) September 29, 2025
✨ Built on V3.1-Terminus, it debuts DeepSeek Sparse Attention(DSA) for faster, more efficient training & inference on long context.
👉 Now live on App, Web, and API.
💰 API prices cut by 50%+!
1/n
9月30日 GLM-4.6
11月6日 Kimi K2 Thinking
11月25日 Sarashina2.2-Vision-3B
12月1日 DeepSeek-V3.2 ・ DeepSeek-V3.2-Speciale
12月2日 Mistral 3
12月9日 Devstral 2
12月17日 MiMo-V2-Flash
2026年
1月5日 LFM 2.5-JP ・ LFM 2.5-VL
1月16日 ELYZA-LLM-Diffusion
1月19日 GLM-4.7 Flash
1月22日 NVIDIA PersonaPlex
1月26日 Kimi K2.5
2月12日 GLM-5
2月16日 Qwen3.5-397B-A17B
2月17日 Nemotron-Nano-9B-v2-Japanese
2月20日 Qwen3 Swallow
2月20日 GPT-OSS Swallow
2月24日 Qwen3.5-122B-A10B ・ Qwen3.5-35B-A3B ・ Qwen3.5-27B
2月24日 Mercury 2
2月25日 LLM-jp-Moshi-v1
2月28日 SSE: Stable Static Embedding 日本語版
3月2日 Qwen3.5-0.8B · Qwen3.5-2B · Qwen3.5-4B · Qwen3.5-9B
3月11日 NVIDIA-Nemotron-3-Super-120B-A12B
3月17日 RakutenAI-3.0
3月17日 Mistral Small 4
3月19日 Nemotron-Cascade 2
3月20日 Nemotron 3 Content Safety 4B
3月24日 hf-mount
3月27日 ABEJA-Qwen3-14B-Agentic-256k-v0.1
3月27日 karakuri-vl-2-8b-thinking-2603
3月31日 OneComp
3月31日 Bonsai 8B
4月2日 Gemma 4
4月3日 PLaMo 2.1-VL
4月3日 LLM-jp-4
4月7日 GLM-5.1
4月14日 LLM-jp-4-VL
4月15日 Qwen3.6-35B-A3B
4月16日 Ternary Bonsai
4月22日 Kimi K2.6
4月22日 Qwen3.6-27B
4月23日 DeepSeek-V4-Pro ・ DeepSeek-V4-Flash
4月28日 Nemotron 3 Nano Omni
4月29日 Mistral Medium 3.5
5月20日 Cohere Command A+
6月1日 MiniMax M3
6月3日 Gemma 4 12B
6月4日 Nemotron 3 Ultra
6月5日 Gemma 4 QAT
6月6日 LFM2.5-Audio-1.5B-JP ・ LFM2.5-1.2B-JP-202606
6月9日 North Mini Code
6月12日 Kimi-2.7-Code
6月16日 GLM-5.2
Introducing GLM-5.2: Frontier Intelligence, Open Weights
— Z.ai (@Zai_org) June 16, 2026
- Significant improvements in coding and agentic tasks
- Strong long-horizon capabilities with a 1M context window
- Two levels of reasoning effort: GLM-5.2 (max) pushes the limits, while GLM-5.2 (high) strikes a strong… pic.twitter.com/SjGPSVhePJ
7月7日 Hy3
🚀Hy3 is here.
— Tencent Hy (@TencentHunyuan) July 6, 2026
295B MoE. Best in its size class. Rivals trillion-scale flagships.
Reliable and affordable for most agentic usecases.
Apache 2.0. Friendly for commercial use.
FREE API for 2 weeks → https://t.co/EyURKwTdgi
🤗 https://t.co/twqJpqb2SL
📖 https://t.co/4uEkIU1cW4 pic.twitter.com/LLJTfDFNQu
7月14日 Bonsai 27B
7月14日 Agents A1 4B
7月15日 Inkling
7月17日 Nemotron 3 Embed 8B
7月27日 Kimi K3
7月31日 DeepSeek-V4-Flash-0731
8月4日 LFM2.5-2.6B
8月10日 Muse Glimmer
8月11日 NVIDIA Nemotron 3.5 Lightning
8月12日 Qwen3.8-2.4T-A95B
8月13日 DeepSeek-V4-Pro-0813
8月14日 Qwen3.8-27B
We promised open weights for Qwen3.8. Now, time to meet them! 🎉
— Qwen (@Alibaba_Qwen) August 14, 2026
⚡ Qwen3.8-27B:
- A native multimodal dense model. With just 27B parameters, it outperforms Qwen3.7-Plus overall and shines in real-world coding & office workflows.
- 262K native context, easily extendable to 1M… pic.twitter.com/QuN8oWkG4C
