CUSTOM KNOWLEDGE FEED

#benchmarks

11 cards

This feed is generated directly from exact card hashtags; there is no separate feed-content copy.

Subscribe to this viewRSSJSON

Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on Large Language Models and exploring the boundaries and limits of Generative AI. Use it to build a structured path from fundamentals to hands-on practice.

MARKDOWN SNAPSHOT

Loading…

00

Добавить первичную статистику + methodology на 10 страницах и сравнить citation absorption. Comparison table vs narrative-only на matched pages.

MARKDOWN SNAPSHOT

Loading…

00

MedAlpaca expands upon both Stanford Alpaca and AlpacaLoRA to offer an advanced suite of large language models specifically fine-tuned for medical question-answering and dialogue applications. Our primary objective is to deliver an array of open-source language models, paving the way for seamless development of medical chatbot solutions.

MARKDOWN SNAPSHOT

Loading…

00

This repository is the reading list on Deep Learning for Mathematical Reasoning (DL4MATH). Contributors: Pan Lu @UCLA, Liang Qiu @UCLA, Wenhao Yu @Notre Dame, Sean Welleck @UW, Kai-Wei Chang @UCLA For more details, please refer to the paper: A Survey of Deep Learning for Mathematical Reasoning.

MARKDOWN SNAPSHOT

Loading…

00