Skip to content

Identify and Summarize Suitable Benchmark Datasets #36

Description

@suppathak

We need to identify benchmark datasets that can be used to evaluate tool selection accuracy, scalability, and latency for LlamaStack. These datasets should include a diverse set of queries requiring various tools.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Type

No type

Projects

No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions