This repository provides the code for fine-tuning BioBERT, a biomedical language representation model designed for biomedical text mining tasks such as biomedical named entity recognition, relation extraction, question answering, etc. Please refer to our paper BioBERT: a pre-trained biomedical language representation model for biomedical text mining for more details.
Welcome to the Cleaned Alpaca Dataset repository! This repository hosts a cleaned and curated version of a dataset used to train the Alpaca LLM (Large Language Model).
pg-mem is an experimental in-memory emulation of a postgres database. ⭐ this repo if you like this package, it helps to motivate me :) 👉 See it in action with pg-mem playground As always, it starts with an: Then, assuming you're using something like webpack, if you're targeting a browser: Pretty straightforward :) ❤ Head to the pgsql-ast-parser repo The sql syntax parser is home-made.
A quick reminder of all relevant SQL queries and examples on how to use them. This repository is constantly being updated and added to by the community.
Walrus: A Distributed Message Streaming Engine Walrus is a distributed message streaming platform built on a high-performance log storage engine. It provides fault-tolerant streaming with automatic leadership rotation, segment-based partitioning, and Raft consensus for metadata coordination.
This README gives an overview of how to build and contribute to the documentation of Apache Flink. The documentation is included with the source of Apache Flink in order to ensure that you always have docs corresponding to your checked out version.
Easy to use open source fast database for search Manticore Search is an easy-to-use, open-source, and fast database designed for search. Website • Downloads • Docs • Blog • Courses • Forum • Slack • Telegram (En) • Telegram (Ru) • Twitter • User feedback What distinguishes Manticore from other solutions is: Results are workload-specific; use the linked dashboard to select the queries that match your workload.
Đọc bản tiếng Việt A complete SQL (and also non-SQL) databases of Vietnamese administrative units, includes all 34 Vietnamese provinces and associated districts, wards sub-divisions. Data is updated as of the most recent effective decree: [30/2026/QH16][source government decree] Add-on includes: GIS Dataset If you find this repository helpful, please consider giving it a ⭐ — it helps us stay motivated to keep improving and delivering valuable tools for the community.
Docs || Official Website || Research Paper English || 简体中文 Connect with us: Contents ======== MatrixOne is the industry's first database to bring Git-style version control to data, combined with MySQL compatibility, AI-native capabilities, and cloud-native architecture. At its core, MatrixOne is a HTAP (Hybrid Transactional/Analytical Processing) database with a hyper-converged HSTAP engine that seamlessly handles transactional (OLTP), analytical (OLAP), full-text search, and