Loading the SOTA2 catalog…
Light Alignment Improves LLM Safety via Model Self-Reflection with a Single Neuron · SOTA2 Research