Skip to content
#

aisafety

Here are 75 public repositories matching this topic...

Materials for the course Principles of AI: LLMs at UPenn (Stat 9911, Spring 2025). LLM architectures, training paradigms (pre- and post-training, alignment), test-time computation, reasoning, safety and robustness (jailbreaking, oversight, uncertainty), representations, interpretability (circuits), etc.

  • Updated Jun 14, 2025

Add this topic to your repo

To associate your repository with the aisafety topic, visit your repo's landing page and select "manage topics."

Learn more