PGCodeLLM
Popular repositories Loading
-
trl
trl PublicForked from huggingface/trl
[Downstream Fork DO NOT EDIT MAIN] Train transformer language models with reinforcement learning.
Python
-
OpenRLHF
OpenRLHF Public[Fork] An Easy-to-use, Scalable and High-performance RLHF Framework (70B+ PPO Full Tuning & Iterative DPO & LoRA & RingAttention & RFT)
Python
-
pytest-json-report
pytest-json-report PublicForked from numirias/pytest-json-report
🗒️ A pytest plugin to report test results as JSON
Python
-
critic-rl
critic-rl PublicForked from HKUNLP/critic-rl
Code for Paper: Teaching Language Models to Critique via Reinforcement Learning
Python
-
-
LLaMA-Factory
LLaMA-Factory PublicForked from hiyouga/LlamaFactory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Python
Repositories
- harbor Public Forked from harbor-framework/harbor
Harbor is a framework for running agent evaluations and creating and using RL environments.
- SWE-gen Public Forked from abundant-ai/SWE-gen
Convert GitHub PRs into Harbor tasks. Now with OBS cloning and Voyager^tm uploading capabilities!
- dsm-ae Public
DSM-AE: diagnostic engine for agentic ill-behaviours (indicator protocols + multi-model matrix)
- SWE-Interact Public Forked from scaleapi/SWE-Interact
New testbed of interactive SWE tasks for coding agents, set in a realistic multi-turn developer driven environment
- emnlp-rebuttal Public
- greenfield Public Forked from prime-radiant-inc/greenfield
A Claude Code plugin that reverse-engineers clean behavioral specs, test vectors, and acceptance criteria from any codebase, producing a provenance trail so a fresh team can reimplement without inheriting the original's internal structure.
- throw-away-repo2 Public
- slop-code-bench Public Forked from SprocketLab/slop-code-bench
SlopCodeBench: Measuring Code Erosion Under Iterative Specification Refinement
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…