-
Notifications
You must be signed in to change notification settings - Fork 3
All issues
Issue creation is restricted in this repository
Issues
is:issue state:open
is:issue state:open
Search results
[Task/Feature/Issue] V1.5 of the TSEI reorganization proposal
documentationImprovements or additions to documentationImprovements or additions to documentationStatus: Open.#58 In The-AI-Alliance/trust-safety-evals;[Task/Feature/Issue] Integrate the Infosys Responsible AI Toolkit
reference stackAll tools for the reference stack.All tools for the reference stack.Status: Open.#56 In The-AI-Alliance/trust-safety-evals;[Docs/Website] Provide some teaching examples
documentationImprovements or additions to documentationImprovements or additions to documentationStatus: Open.#55 In The-AI-Alliance/trust-safety-evals;[Task/Feature/Issue] Add MedHelm
enhancementNew feature or requestNew feature or requestevaluationsEvaluations and their implementations, including suites of evaluations that make up benchmarks.Evaluations and their implementations, including suites of evaluations that make up benchmarks.leaderboardsLeaderboards deployed to HF or other placesLeaderboards deployed to HF or other placestaxonomyThe definition and review tasks for the global taxomonyThe definition and review tasks for the global taxomonyStatus: Open.#50 In The-AI-Alliance/trust-safety-evals;Investigate the Credo AI graphical browser.
taxonomyThe definition and review tasks for the global taxomonyThe definition and review tasks for the global taxomonyStatus: Open.#48 In The-AI-Alliance/trust-safety-evals;- Status: Open.#41 In The-AI-Alliance/trust-safety-evals;
- Status: Open.#40 In The-AI-Alliance/trust-safety-evals;
Create an example using Arize Phoenix
reference stackAll tools for the reference stack.All tools for the reference stack.Status: Open.#39 In The-AI-Alliance/trust-safety-evals;Incorporate Databricks "Domain Intelligence" benchmark
evaluationsEvaluations and their implementations, including suites of evaluations that make up benchmarks.Evaluations and their implementations, including suites of evaluations that make up benchmarks.taxonomyThe definition and review tasks for the global taxomonyThe definition and review tasks for the global taxomonyStatus: Open.#38 In The-AI-Alliance/trust-safety-evals;- Status: Open.
Evaluate LangFair as a tool for evaluations and ideas for the taxonomy
evaluationsEvaluations and their implementations, including suites of evaluations that make up benchmarks.Evaluations and their implementations, including suites of evaluations that make up benchmarks.help wantedExtra attention is neededExtra attention is neededtaxonomyThe definition and review tasks for the global taxomonyThe definition and review tasks for the global taxomonyStatus: Open.Talk with MLCommons about incorporating their evaluations, benchmarks, etc.
collaborators"Strategic" work with third-party collaborators"Strategic" work with third-party collaboratorsevaluationsEvaluations and their implementations, including suites of evaluations that make up benchmarks.Evaluations and their implementations, including suites of evaluations that make up benchmarks.Status: Open.#31 In The-AI-Alliance/trust-safety-evals;