This is a comparison of various models that are evauluated on the same log set. We are generting the full Gitlab MR comment with them.
Right now we have results here from 2 models:
Make sure that logdetective.server is available on your $PYTHONPATH because
the run.py script is using functions from it. For gemini I had to do 2 code changes:
- Comment out all logprobs operations because gemini doesn't implement logprobs
- Drop the
/v1/prefix in the URL
$ python3 run.py local-config.yaml $dir_with_results
Arguments:
- logdetective server config, one is included in the repo, gemini config is below; make sure your inference is running
- dir where the results should be put (default current working dir)
Sample gemini config:
log:
level_stream: "DEBUG"
level_file: "DEBUG"
format: "%(asctime)s - %(levelname)s - %(message)s"
inference:
max_tokens: 10000
# log_probs: 1 this doesn't work with gemini openai api
url: https://generativelanguage.googleapis.com/v1beta/openai
api_token: abcdef
model: gemini-2.5-flash-preview-04-17 # gemini-2.5-flash
extractor:
context: true
max_clusters: 25
verbose: false