| 70PyMuPDF4LLM | Library that makes it easier to extract PDF content in the format you need for LLM & RAG environments. | LLM Data Extraction | https://github.com/PyMuPDF4LLM/PyMuPDF4LLM |
|---|
| 71Crawlee | Web scraping and browser automation library. | LLM Data Extraction | https://github.com/apify/crawlee |
|---|
| 72MegaParse | Parser for every type of document. | LLM Data Extraction | https://github.com/MegaParse/MegaParse |
|---|
| 73ExtractThinker | Document intelligence library for LLMs. | LLM Data Extraction | https://github.com/ExtractThinker/ExtractThinker |
|---|
| 74DataDreamer | DataDreamer is a powerful open-source Python library for prompting, synthetic data generation, and training workflows. | LLM Data Generation | https://github.com/datadreamer-dev/DataDreamer |
|---|
| 75Fabricator | A flexible open-source framework to generate datasets with large language models. | LLM Data Generation | https://github.com/eyurtsev/fabricator |
|---|
| 76Promptwright | Synthetic Dataset Generation Library. | LLM Data Generation | https://github.com/promptwright/promptwright |
|---|
| 77EasyInstruct | An Easy-to-use Instruction Processing Framework for Large Language Models. | LLM Data Generation | https://github.com/zjunlp/EasyInstruct |
|---|
| 78CrewAI | Framework for orchestrating role-playing, autonomous AI agents. | LLM Agents | https://github.com/crew-ai/crew |
|---|
| 79LangGraph | Build resilient language agents as graphs. | LLM Agents | https://github.com/langgraph/langgraph |
|---|
| 153 more items in this listSign up free to see them all, with every column and every photograph.Sign up free |