Skip to content
View pr-Mais's full-sized avatar
👾
👾

Organizations

@googlemaps @fluttercommunity @FlutterVikings @Thmanyah-LLC

Block or report pr-Mais

Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Curated list of Go design patterns, recipes and idioms

Go 28,256 2,338 Updated May 14, 2024

Terminal UI for AWS (taws) - A terminal-based AWS resource viewer and manager

Rust 2,261 71 Updated Sep 27, 2026

🐹 Clean, uninstall, analyze, optimize, and monitor your Mac. Free open-source CLI, plus a native Mac app.

Shell 68,898 2,443 Updated Sep 30, 2026

ASU-sparkysundevil-resume-template

TeX 36 19 Updated Oct 3, 2024

The conventional commits specification

SCSS 9,283 696 Updated Mar 11, 2026

A First Look at Conventional Commits Classification

Python 16 1 Updated Nov 18, 2024

NPM library to splice HLS VOD

JavaScript 19 5 Updated Sep 10, 2026

💻 A fully functional local AWS cloud stack. Develop and test your cloud & Serverless apps offline

Python 65,133 4,807 Updated Mar 23, 2026

methods2test is a supervised dataset consisting of Test Cases and their corresponding Focal Methods from a set of Java software repositories

Python 172 48 Updated Dec 4, 2023

Analysis scripts for log data sets used in anomaly detection.

Python 85 17 Updated Oct 19, 2025

Firebase SDK for Cloud Functions

TypeScript 1,069 233 Updated Sep 17, 2026

Implementation of ChatGPT RLHF (Reinforcement Learning with Human Feedback) on any generation model in huggingface's transformer (blommz-176B/bloom/gpt/bart/T5/MetaICL)

Python 565 61 Updated Apr 23, 2026

Mastering Diverse Domains through World Models

Python 3,840 621 Updated May 25, 2026

Train transformer language models with reinforcement learning.

Python 19,428 3,035 Updated Oct 1, 2026

Schedule-Free Optimization in PyTorch

Python 2,325 79 Updated Jul 28, 2026

Fine-tune LLM agents with online reinforcement learning

Python 1,253 65 Updated Mar 19, 2024

A library with extensible implementations of DPO, KTO, PPO, ORPO, and other human-aware loss functions (HALOs).

Python 909 51 Updated Sep 30, 2025

Reference implementation for DPO (Direct Preference Optimization)

Python 2,910 235 Updated Aug 11, 2024

FastAPI framework, high performance, easy to learn, fast to code, ready for production

Python 102,745 9,977 Updated Oct 1, 2026

PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.

Python 13,862 2,182 Updated Sep 9, 2026

Powerful menu bar manager for macOS

Swift 29,730 966 Updated Sep 20, 2025
Python 3 Updated Sep 16, 2026

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

Python 7,860 672 Updated Sep 20, 2026

Implementation of the ICML 2024 paper "Training Large Language Models for Reasoning through Reverse Curriculum Reinforcement Learning" presented by Zhiheng Xi et al.

Python 117 10 Updated Feb 9, 2024

Comprehensive toolkit for Reinforcement Learning from Human Feedback (RLHF) training, featuring instruction fine-tuning, reward model training, and support for PPO and DPO algorithms with various c…

Python 194 19 Updated Sep 3, 2026

Monitoring recent cross-research on LLM & RL on arXiv for control. If there are good papers, PRs are welcome.

562 36 Updated Nov 17, 2025
Python 146 15 Updated May 2, 2024

Curated tutorials and resources for Large Language Models, Text2SQL, Text2DSL、Text2API、Text2Vis and more.

3,761 258 Updated Jan 26, 2026

Data for paper "Dr.Spider: A Diagnostic Evaluation Benchmark towards Text-to-SQL Robustness"

Python 34 7 Updated May 3, 2023
Python 187 51 Updated Sep 5, 2026
Next