Pinned Loading
-
Qwen-0.5B-Reasoning-GRPO
Qwen-0.5B-Reasoning-GRPO PublicCan a 500M parameter model actually "think"? "Inducing multi-step reasoning and CoT in Qwen-0.5B-Instruct via GRPO."
Python
-
Multimodal-GPT2-From-Scratch
Multimodal-GPT2-From-Scratch PublicTeaching a 124M parameter text-only model to read images: A deep-dive into projection-based alignment. No API wrappers, just raw tensor mapping from vision to language latent space
Python
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.