Shielded RecRL

This repository contains the implementation of Shielded RecRL, a method for adding chat-style explanations to recommender systems without affecting the underlying ranking model.

Project Overview

Shielded RecRL uses a two-tower architecture:

A frozen ranking model (collaborative filtering)
A trainable language model that generates explanations

The key innovation is the gradient projection technique that prevents the explanation model from affecting the ranking model's performance.

Setup Instructions

Local Setup (Any OS)

Clone this repository:

git clone https://github.com/your_username/shielded-recrl.git
cd shielded-recrl

Edit setup_local.sh to update your GitHub username, then run:
```
bash setup_local.sh
```

RunPod Setup (Remote GPU)

Launch a RunPod instance with:
- Runtime: PyTorch 2.3 | Python 3.10 | CUDA 12.2
- GPU: NVIDIA A100 80GB or 2× RTX 4090 24GB
- Volume: ≥ 400GB

SSH into your RunPod instance:

ssh -p YOUR_PORT runpod@YOUR_POD_ID.connect.runpod.io

Edit setup_runpod.sh to update your GitHub username, then run:
```
bash setup_runpod.sh
```
Verify the setup:
```
python gpu_test.py
```

Project Structure

├── code
│   ├── dataset/    # Dataset preprocessing
│   ├── ranker/     # SASRec implementation
│   ├── explainer/  # LLM with LoRA
│   ├── projection/ # Gradient projection
│   ├── trainer/    # Shielded PPO
│   └── eval/       # Evaluation metrics
├── data           # Datasets
├── checkpoints    # Model checkpoints
├── logs           # Training logs
├── experiments    # Experiment configurations
├── docs           # Documentation
└── docker         # Docker configuration

Workflow

Edit code on your local machine
Commit and push changes to GitHub
Pull changes on RunPod and execute experiments
Results are logged to W&B and saved to the persistent volume

License

[Add your license information here]

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Shielded RecRL

Project Overview

Setup Instructions

Local Setup (Any OS)

RunPod Setup (Remote GPU)

Project Structure

Workflow

License

About

Uh oh!

Releases

Packages

Uh oh!

Contributors

Uh oh!

Languages

Name		Name	Last commit message	Last commit date
Latest commit History 1 Commit
.github/workflows		.github/workflows
code		code
data		data
docker		docker
docs		docs
experiments		experiments
.gitignore		.gitignore
README.md		README.md
environment_check.sh		environment_check.sh
gpu_test.py		gpu_test.py
requirements.txt		requirements.txt
setup_local.sh		setup_local.sh
setup_runpod.sh		setup_runpod.sh

Folders and files

Latest commit

History

Repository files navigation

Shielded RecRL

Project Overview

Setup Instructions

Local Setup (Any OS)

RunPod Setup (Remote GPU)

Project Structure

Workflow

License

About

Resources

Uh oh!

Stars

Watchers

Forks

Releases

Packages 0

Uh oh!

Contributors

Uh oh!

Languages

Packages