About the challenge
Over a single six-hour session, your group will build a safety evaluation report using Gemma LLMs (gemma-4-E2B and gemma-4-E2B-it). Each group selects the domain/use-case they will target. All code implementations will be done during the hackathon. The project will focus on safety and involve understanding the limitations of transformer model architecture.
Requirements
What to Build
Over a single six-hour session, your group will build a safety evaluation scorecard for an LLM. Each group selects the domain and the metrics they want to include in their scorecard, then implements and runs those metrics against a provided model. For the purposes of this hackathon, you will be acting as if the general-purpose LLM we provide you was created for the purpose of your use-case.
Groups will create a complete pitch for a safety scorecard — all the metrics that should be included and why, specific to your chosen domain and use-case — and implement them during the hackathon. The main goal is to grapple with three questions: What should we measure, Why should we measure it, and How can we measure it? Approach each from a model architecture standpoint and note the limitations inherent in every measurement.
This is an evaluation exercise, not a training exercise. You are not improving, fine-tuning, or otherwise training the model in any way. The work is figuring out how to correctly evaluate the necessary safety dimensions for your use-case and recognize the limitations of those measurements given by the architecture of current-day LLMs.
What to Submit
- Project Overview
- Code on GitHub
- A table of your results
- Written Discussion
- Brief list of what each individual contributed
- ~5 min Video
Prizes
1st Place
You are the overall winner of this hackathon
2nd Place
You are the runner up of this hackathon
3rd Place
You are the 3rd place winner of this hackathon
Devpost Achievements
Submitting to this hackathon could earn you:
Judges
Eileanor LaRocco
Chirag Agarwal
Judging Criteria
-
Innovation and Creativity (20 pts)
Is this a new idea or unique approach to an old problem? -
Technical Implementation (30 pts)
Does the prototype work? -
Problem-Solving & Impact (20 pts)
Does this project solve a real, meaningful problem in AI safety? Are the choices and limitations clearly discussed? -
Presentation/Pitch (10 pts)
Presentation of your results on Devpost!
Questions? Email the hackathon manager
Tell your friends
This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.