Highlights
- Pro
Pinned Loading
-
-
ellanorai/ThinkRL
ellanorai/ThinkRL PublicThinkRL is a comprehensive, state-of-the-art reinforcement learning from human feedback (RLHF) library designed to democratize advanced AI training. Built with a zero-dependency core philosophy, Th…
-
-
-
-
smolagent-ollama
smolagent-ollama PublicA AI Agent with smolagent library using local ollama'l llama3.1:latest
Python
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.

