Best Productivity AI Tools

1,173 tools

OneKE logo

OneKE

A bilingual Chinese-English knowledge extraction model with knowledge graphs and natural language processing technologie

Freemium
PromptLayer 🍰 logo

PromptLayer 🍰

Prompt Engineering platform. Collaborate, test, evaluate, and monitor your LLM applications

Freemium
Puzzlet AI logo

Puzzlet AI

The Git-Based LLM Engineering Platform. Achieve more from GenAI: Manage, evaluate, and improve your full-stack LLM appli

Freemium
PromptHub logo

PromptHub

Full stack prompt management tool designed to be usable by technical and non-technical team members. Test, version, coll

Freemium
MovieLens-1M logo

MovieLens-1M

dataset, embodying varied social traits and preferences.

Freemium
AlpacaEval logo

AlpacaEval

An Automatic Evaluator for Instruction-following Language Models using Nous benchmark suite.

Freemium
Parea AI logo

Parea AI

Platform and SDK for AI Engineers providing tools for LLM evaluation, observability, and a version-controlled enhanced p

Freemium
Manag.ai logo

Manag.ai

Your all-in-one prompt management and observability platform. Craft, track, and perfect your LLM prompts with ease.

Freemium
Izlo logo

Izlo

Prompt management tools for teams. Store, improve, test, and deploy your prompts in one unified workspace.

Freemium
Berkeley Function-Calling Leaderboard logo

Berkeley Function-Calling Leaderboard

evaluates LLM's ability to call external functions/tools.

Freemium
CompMix logo

CompMix

a benchmark evaluating QA methods that operate over a mixture of heterogeneous input sources (KB, text, tables, infoboxe

Freemium
Epsilla logo

Epsilla

An all-in-one platform to create vertical AI agents powered by your private data and knowledge.

Freemium
FELM logo

FELM

a meta-benchmark that evaluates how well factuality evaluators assess the outputs of large language models (LLMs).

Freemium
LLMEval logo

LLMEval

focuses on understanding how these models perform in various scenarios and analyzing results from an interpretability pe

Freemium
MathEval logo

MathEval

a comprehensive benchmarking platform designed to evaluate large models' mathematical abilities across 20 fields and nea

Freemium
MMedBench logo

MMedBench

a benchmark that evaluates large language models' ability to answer medical questions across multiple languages.

Freemium
LLMSTXT.NEW logo

LLMSTXT.NEW

Generate consolidated text files from websites for LLM training and inference – Powered by Firecrawl

Freemium
Marvin logo

Marvin

AI engineering framework for building natural language interfaces

Freemium
LMExamQA logo

LMExamQA

a leaderboard that benchmarks foundation models with Language-Model-as-an-Examiner.

Freemium
LLM Use Case Leaderboard logo

LLM Use Case Leaderboard

a leaderboard that features LLM use cases.

Freemium
Evaluating LLMs is a minefield logo

Evaluating LLMs is a minefield

talk by Princeton professor Arvind Narayanan

Freemium
LiveBench logo

LiveBench

A Challenging, Contamination-Free LLM Benchmark.

Free
Eden AI logo

Eden AI

provides a unique API connected to the AI engines

Freemium
ChatArena logo

ChatArena

building multi-agent environments for LLMs

Freemium