InternLM

InternLM

9 models in Artifacts Logshuggingface.co/internlm
Intern-S2-Preview-397B16 Jul 2026
397BApache-2.0420 · 30d

A multimodal model aimed at scientific applications.

Intern-S1-Pro02 Feb 2026
1T Total (22B Active)Apache-2.0369.4K · 30d

A 1T model by InternLM/Shanghai AI Laboratory, which focuses on STEM. However, its real utility seems to be below what the benchmarks suggest.

Intern-S124 Jul 2025
241BApache-2.08K · 30d

One of the first multimodal Qwen3 fine-tunes. This release from the talented InternLM team combines the large Qwen3 MoE with their own ViT.

OREAL-32B10 Feb 2025
32BApache-2.065 · 30d

A bit of a different take on the current rage of reasoning models. Quoting the model card: "Our method leverages best-of-N (BoN) sampling for behavior cloning and reshapes negative sample rewards to ensure gradient consistency. Also, to address the challenge of sparse rewards in long chain-of-thought reasoning, we incorporate an on-policy token-level reward model that identifies key tokens in reasoning trajectories for importance sampling." There are more details in the paper, but mostly this goes to show that reward models (specifically, outcome reward models) are still important to reasoning and it isn't that "explicit verification is all you need."

OREAL-7B10 Feb 2025
7BApache-2.064 · 30d

A bit of a different take on the current rage of reasoning models. Quoting the model card: "Our method leverages best-of-N (BoN) sampling for behavior cloning and reshapes negative sample rewards to ensure gradient consistency. Also, to address the challenge of sparse rewards in long chain-of-thought reasoning, we incorporate an on-policy token-level reward model that identifies key tokens in reasoning trajectories for importance sampling." There are more details in the paper, but mostly this goes to show that reward models (specifically, outcome reward models) are still important to reasoning and it isn't that "explicit verification is all you need."

InternLM-XComposer2.5-Reward21 Jan 2025
7BOther109 · 30d

A multi-modal reward model.

InternLM3-8B-Instruct13 Jan 2025
8BApache-2.086.7K · 30d

Similar to the other Chinese labs, InternLM has also updated their model series before Chinese New Year.

internlm2-20b-reward27 Jun 2024
104 · 30d

Another very strong reward model on RewardBench. Trained on 3.5 million preference pairs, which is way more than most. A good trend.

internlm2-math-plus-mixtral8x22b24 May 2024
41 · 30d

Next model in the popular series of math models.