Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Directory
Search results
Published directory entries matching your search.
An English (Porter2) stemming implementation in Elixir.
A library for processing Chinese text.
Natural Language Understanding library for intent classification and entity extraction.
A C++ library for unsupervised text tokenization and detokenization, widely used in modern NLP models.
Chrome extension that uses local LLMs to assist with writing and drafting responses based on the context of your open tabs.
Simple and efficient library to minimize expensive and noisy black-box functions.
A tool to help you configure, organize, log and reproduce experiments.
Fast MATLAB-syntax runtime with automatic CPU/GPU execution and fused array kernels.
Text processing tools and wrappers (e.g. Vowpal Wabbit).
Extensible system for analyzing and manipulating natural language.
A "machine learning framework to automate text-and voice-based conversations.".
Statistics, data mining and machine learning toolbox in Java.
CLI tool for debugging and benchmarking RAG retrieval. EXPLAIN ANALYZE for your retrieval layer.
File Parser optimised for LLM Ingestion with no loss. Parse PDFs, Docx, PPTx in a format that is ideal for LLMs.
A Python extension module wrapping the full TiMBL C++ programming interface. Timbl is an elaborate k-Nearest Neighbours machine learning toolkit.
Multilingual text (NLP) processing toolkit.
A platform for reproducible and scalable machine learning and deep learning on kubernetes.
Python Machine Learning Pentesting Toolbox for Adversarial Attacks. Works with LLMs, DNNs, and other machine learning algorithms.
Retrieval Augmented Generation (RAG) framework and context engine powered by Pinecone.
A complete object-oriented environment for machine learning in Matlab.
Lambda Architecture Framework using Apache Spark and Apache Kafka with a specialization for real-time large-scale machine learning.
Gitingest - Turn any Git repository into a simple text digest of its codebase so it can be fed into any LLM.
Source code and experiments results for Santander Value Prediction Challenge.