About the challenge

Over a single six-hour session, your group will build a safety evaluation report using Gemma LLMs (gemma-4-E2B and gemma-4-E2B-it). Each group selects the domain/use-case they will target. All code implementations will be done during the hackathon. The project will focus on safety and involve understanding the limitations of transformer model architecture.

DETAILED INSTRUCTIONS

 

Requirements

What to Build

Over a single six-hour session, your group will build a safety evaluation scorecard for an LLM. Each group selects the domain and the metrics they want to include in their scorecard, then implements and runs those metrics against a provided model. For the purposes of this hackathon, you will be acting as if the general-purpose LLM we provide you was created for the purpose of your use-case. 

Groups will create a complete pitch for a safety scorecard — all the metrics that should be included and why, specific to your chosen domain and use-case — and implement them during the hackathon. The main goal is to grapple with three questions: What should we measure, Why should we measure it, and How can we measure it? Approach each from a model architecture standpoint and note the limitations inherent in every measurement. 

This is an evaluation exercise, not a training exercise. You are not improving, fine-tuning, or otherwise training the model in any way. The work is figuring out how to correctly evaluate the necessary safety dimensions for your use-case and recognize the limitations of those measurements given by the architecture of current-day LLMs.

What to Submit

  • Project Overview
  • Code on GitHub
  • A table of your results
  • Written Discussion
  • Brief list of what each individual contributed
  • ~5 min Video

Hackathon Sponsors

Prizes

3 non-cash prizes
1st Place
1 winner

You are the overall winner of this hackathon

2nd Place
1 winner

You are the runner up of this hackathon

3rd Place
1 winner

You are the 3rd place winner of this hackathon

Devpost Achievements

Submitting to this hackathon could earn you:

Judges

Eileanor LaRocco

Eileanor LaRocco

Chirag Agarwal

Chirag Agarwal

Judging Criteria

  • Innovation and Creativity (20 pts)
    Is this a new idea or unique approach to an old problem?
  • Technical Implementation (30 pts)
    Does the prototype work?
  • Problem-Solving & Impact (20 pts)
    Does this project solve a real, meaningful problem in AI safety? Are the choices and limitations clearly discussed?
  • Presentation/Pitch (10 pts)
    Presentation of your results on Devpost!

Questions? Email the hackathon manager

Tell your friends

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.