Skip to content
View jacobphillips99's full-sized avatar

Block or report jacobphillips99

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Please don't include any personal information such as legal names or email addresses. Maximum 250 characters, markdown supported. This note will be visible to only you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. open-rubric open-rubric Public

    Forked from PrimeIntellect-ai/verifiers

    Fork of verifiers focused on multi-step rubric evaluation complete with multi-step environments and synthetic data generators

    Python 4

  2. ares ares Public

    ARES - Automatic Robot Evaluation System. A simple, scalable solution for robotics research

    Python 55 8

  3. daily-bench daily-bench Public

    A daily benchmark to regression-test cloud LLMs

    JavaScript 11 2

  4. llm-rate-limiter llm-rate-limiter Public

    Simple rate limiter for LLM API calls

    Python

  5. mallet mallet Public

    Cloud-based tools and an evaluation harness for VLMs to control real-world robots

    Python 9

  6. protorubric protorubric Public

    Custom agent framework for multi-step LLM-as-Judge rubric evaluation

    Python