Skip to content

Need Complete Example for Running Evaluation #10

Description

@AOoligei

Hi! The current example (local_BGE_local_LLM.py) is too simplified - it only shows a basic query without any evaluation.

Could you provide a complete example that shows:

  • How to load and evaluate on actual datasets (e.g., ComplexTR)
  • How to calculate metrics (accuracy, F1, etc.)

Basically, how do I reproduce the paper's results?

Thanks!

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions