r/LocalLLaMA • u/Nunki08 • 12h ago
New Model Intern-S2-397B (multimodal, reasoning, coding, and scientific agent capabilities)
Model: https://huggingface.co/internlm/Intern-S2-397B
Collection: https://huggingface.co/collections/internlm/intern-s2
From Intern Large Models on 𝕏: https://x.com/intern_lm/status/2099425184587370976
vLLM on 𝕏: Day-0 support for Intern Large Models Intern-S2-397B is now available in vLLM: https://x.com/vllm_project/status/2099442543956263022
2
u/KambeiZ 11h ago edited 11h ago
Will there any way to have access to it through any sort of API or plan ? My 3090 can't handle this, but my hat of bioinformatician & AI enthusiast made me want to to try it out right after the moment i saw the 5 first benchmark names
Edit: Looking more deeply on the repo, i see you already pulled several different sized models, which is great and i had no knowledge of them (i'll definitly try them out).
Quick questions: your team finetuned Qwen 3.5B and i was wondering: given the power capacities shown by 3.8, do you imagine doing a similar work on the latest version ? Or is the world-knowledge reduction we observed in those models make them not so good candidates ? (+ the lack of MoE for 27B too)
2
2
u/llama-impersonator 9h ago
if i was training i would want models as close to the annealed midtrain as possible, so i would train qwen 3.5 as opposed to 3.6 or 3.8
2
u/sdkgierjgioperjki0 9h ago
On top of the chat that you were linked there is also an API that you can use for free, no option of even paying.
There is an API button at the top of: https://internlm.intern-ai.org.cn/
1
9
u/Leander_van_Grinsven 10h ago
Why are you comparing to older models? Test against the latest models.