A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
- Stars
- 300
- Forks
- 17
- Pushed
- 10mo ago
Current GitHub adoption and maintenance signals.
A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
[TACL 2024] MAPS enables LLMs🤖 to mimic the human😁 translation process.
[ACL 2024] Can Watermarks Survive Translation? On the Cross-lingual Consistency of Text Watermark for Large Language Models
[ACL 2022] Bridging the Data Gap between Training and Inference for Unsupervised Neural Machine Translation
[WMT 2022] Implementation of TAL-SJTU's system for WMT22 English-Livonian
Code of "Improving Machine Translation with Human Feedback: An Exploration of Quality Estimation as a Reward Model"
{DeepL, Google, WMT-Best, davinci-003, turbo, gpt-4} × {En-De, En-Cs, En-Ru, En-Zh, De-Fr, En-Ja, Uk-En, Uk-Cs, En-Hr, En-Ha, En-Is}
[ICLR 2025] RaSA: Rank-Sharing Low-Rank Adaptation
Some useful scripts to process monolingual or parallel texts.
No repository description provided.
《动手学大模型Dive into LLMs》系列编程实践教程
veRL: Volcano Engine Reinforcement Learning for LLM
Implementation of string group-by algorithm based on radix sort and parallelized using openmp,MPI and MPI+Openmp.
AcadHomepage: A Modern and Responsive Academic Personal Homepage
No repository description provided.
No repository description provided.
The implementation of bitcoin-based timed-commitment scheme from "Secure Multiparty Computations on Bitcoin"
VeOmni: Scaling Any Modality Model Training with Model-Centric Distributed Recipe Zoo
Post-training with Tinker
Official repository for the paper "LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code"
OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.
An Efficient and User-Friendly Scaling Library for Reinforcement Learning with Large Language Models
Awesome RL Reasoning Recipes ("Triple R")
An Easy-to-use, Scalable and High-performance RLHF Framework (70B+ PPO Full Tuning & Iterative DPO & LoRA & RingAttention & RFT)
Train transformer language models with reinforcement learning.
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
A high-throughput and memory-efficient inference and serving engine for LLMs
Tools for merging pretrained large language models.
RewardBench: the first evaluation tool for reward models.
A framework for few-shot evaluation of language models.
A framework for the evaluation of autoregressive code generation language models.
Rigourous evaluation of LLM-synthesized code - NeurIPS 2023
ScienceWorld is a text-based virtual environment centered around accomplishing tasks from the standardized elementary science curriculum.
Fast and memory-efficient exact attention
[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
The lastest paper about detection of LLM-generated text and code
A Neural Framework for MT Evaluation
A Comprehensive Benchmark to Evaluate LLMs as Agents
🐫 CAMEL: Communicative Agents for “Mind” Exploration of Large Scale Language Model Society
Build, evaluate, analyze, and understand LLM-based apps
No repository description provided.
This project collects awesome resources (e.g., papers, open-source models) for large language model (LLM)
The LaTeX beamerposter package
No repository description provided.
This is the homework of C++ Training .