-
Georgia Institute of Technology
- Atlanta, USA
-
-
alpaca_eval Public
Forked from tatsu-lab/alpaca_evalAn automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and fast.
Jupyter Notebook Apache License 2.0 UpdatedDec 27, 2024 -
PASTA Public
PASTA: Post-hoc Attention Steering for LLMs
-
lm-evaluation-harness Public
Forked from EleutherAI/lm-evaluation-harnessA framework for few-shot evaluation of language models.
Python MIT License UpdatedOct 28, 2024 -
transformers Public
Forked from huggingface/transformers🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
Python Apache License 2.0 UpdatedSep 13, 2024 -
OptiMUS Public
Forked from teshnizi/OptiMUSOptimization Modeling Using mip Solvers and large language models
Python MIT License UpdatedSep 4, 2024 -
trl Public
Forked from huggingface/trlTrain transformer language models with reinforcement learning.
Python Apache License 2.0 UpdatedJul 9, 2024 -
OpenRLHF Public
Forked from OpenRLHF/OpenRLHFAn Easy-to-use, Scalable and High-performance RLHF Framework (70B+ PPO Full Tuning & Iterative DPO & LoRA & Mixtral)
Python Apache License 2.0 UpdatedJun 27, 2024 -
-
alignment-handbook Public
Forked from huggingface/alignment-handbookRobust recipes to align language models with human and AI preferences
Python Apache License 2.0 UpdatedMay 25, 2024 -
MAmmoTH Public
Forked from TIGER-AI-Lab/MAmmoTHCode and data for "MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning" (ICLR 2024)
Jupyter Notebook UpdatedApr 25, 2024 -
-
TrustLLM Public
Forked from HowieHwong/TrustLLMTrustLLM: Trustworthiness in Large Language Models
Python MIT License UpdatedMar 14, 2024 -
chain-of-thought-hub Public
Forked from FranxYao/chain-of-thought-hubBenchmarking large language models' complex reasoning ability with chain-of-thought prompting
-
-
peft Public
Forked from huggingface/peft🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
-
math Public
Forked from hendrycks/mathThe MATH Dataset (NeurIPS 2021)
Python MIT License UpdatedDec 4, 2023 -
lost-in-the-middle Public
Forked from nelson-liu/lost-in-the-middleCode and data for "Lost in the Middle: How Language Models Use Long Contexts"
Python MIT License UpdatedNov 20, 2023 -
FastChat Public
Forked from lm-sys/FastChatAn open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
Python Apache License 2.0 UpdatedNov 9, 2023 -
Sophia Public
Forked from Liuhong99/SophiaThe official implementation of “Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training”
Python MIT License UpdatedOct 19, 2023 -
text-generation-webui Public
Forked from oobabooga/textgenA Gradio web UI for Large Language Models. Supports transformers, GPTQ, AWQ, llama.cpp (GGUF), Llama models.
Python GNU Affero General Public License v3.0 UpdatedOct 16, 2023 -
vllm Public
Forked from vllm-project/vllmA high-throughput and memory-efficient inference and serving engine for LLMs
Python Apache License 2.0 UpdatedOct 14, 2023 -
exllama Public
Forked from turboderp/exllamaA more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.
Python MIT License UpdatedSep 30, 2023 -
Data Mining Project: extract raw data from LianJia; data proprocess and feather engineering; build model with MLP and BiLSTM.
-
summarize-from-feedback Public
Forked from openai/summarize-from-feedbackCode for "Learning to summarize from human feedback"
Python Other UpdatedSep 5, 2023 -
AdaLoRA Public
AdaLoRA: Adaptive Budget Allocation for Parameter-Efficient Fine-Tuning (ICLR 2023).
-
unify-parameter-efficient-tuning Public
Forked from jxhe/unify-parameter-efficient-tuningImplementation of paper "Towards a Unified View of Parameter-Efficient Transfer Learning" (ICLR 2022)
Python Apache License 2.0 UpdatedNov 15, 2022 -
PLATON Public
This pytorch package implements PLATON: Pruning Large Transformer Models with Upper Confidence Bound of Weight Importance (ICML 2022).
-
-
baseline-lora Public
Forked from microsoft/LoRACode for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"



