SemanticScuttle - klotz.me » klotz: llm+pdf

klotz: llm* + pdf*

olmOCR: Toolkit for Training Language Models to Work with PDF Documents

A toolkit for training language models to work with PDF documents in the wild, including prompting strategies, evaluation tools, filtering, finetuning code, and processing PDFs through finetuned models.

2025-02-28 Tags: pdf, llm, pdf processing, olmocr, allenai, ocr, document management, document conversion by klotz

Preparing PDFs for RAGs

The article discusses the process of preparing PDFs for use in Retrieval-Augmented Generation (RAG) systems, with a focus on creating graph-based RAGs from annual reports containing tables. It highlights the benefits of Graph RAGs over vector store-backed RAGs, particularly in terms of reasoning capabilities, and explores the construction of knowledge graphs for better information retrieval. The author shares insights into the challenges and solutions involved in building an enterprise-ready graph data store for RAG applications.

2025-01-20 Tags: pdf, rags, knowledge graph, llm by klotz

MarkItDown - Python tool for converting files and office documents to Markdown

MarkItDown is a utility for converting various files to Markdown, including PDF, PowerPoint, Word, Excel, Images, Audio, HTML, text-based formats, and ZIP files.

2024-12-30 Tags: markitdown, markdown, file conversion, python, office documents, pdf, powerpoint, word, excel, images, audio, html, csv, json, xml, zip, openai, large language models, docker, llm, document, conversion by klotz

Improved RAG Document Processing With Markdown

How to read and convert PDFs to Markdown for better RAG results with LLMs.

2024-11-19 Tags: markdown, conversion, pdf, llm by klotz

DS4SD / docling

Docling is a tool that parses documents and exports them to desired formats like Markdown and JSON. It supports various document formats including PDF, DOCX, PPTX, Images, HTML, AsciiDoc, and Markdown.

2024-11-01 Tags: docling, ibm, document, parsing, markdown, json, pdf, docx, pptx, ocr, llm by klotz

Useful LLM Tools

Approximate Tokens, Words and Characters Calculator for LLM's and Text Trimmer — Simple calculator to estimate tokens for Large Language Models and text editor to trim text
Text File Merger for LLM — This tool combines multiple text files into a single document, with clear separation between files
PDF to TXT Converter — Convert PDF documents to plain text format for use with LLMs and text analysis
HTML to TXT Converter — Remove HTML tags and extract clean text content for LLM processing
LLM System Prompt Generator — Generate optimized system prompts for different LLM model sizes (3B, 33B, 70B, etc.)
Creative Idea Generator — AI-powered brainstorming tool for generating creative solutions and ideas

2024-10-26 Tags: llm, tools, linux, denis shiryaev, pdf, html, text by klotz

Scaling ColPali to billions of PDFs with Vespa

This blog post explores scaling ColPali for efficient document retrieval across large collections of PDFs using Vespa's phased retrieval and ranking pipeline, including the use of a hamming-based MaxSim similarity function.

2024-09-23 Tags: colpali, document retrieval, vespa, maxsim, hamming distance, vlm, binary quantization, pdf, vision language models, llm by klotz

IncarnaMind: Chat with Your Documents using LLMs

IncarnaMind enables chatting with personal documents (PDF, TXT) using Large Language Models (LLMs) like GPT. It uses a Sliding Window Chunking mechanism and Ensemble Retriever for efficient querying.

2024-08-09 Tags: llm, pdf, langchain, github by klotz

Improving PDF Parsing and Search with Chunking

A post discussing new techniques developed for parsing and searching PDFs, focusing on turning them into a hierarchical structure for RAG search. The approach involves dynamically generating chunks for searches, sending headers and sub-headers to the Language Model along with relevant chunks.

2024-06-27 Tags: pdf, parsing, search, rag, llm, reddit by klotz

Developer APIs to Accelerate LLM Projects - nlmatics/llmsherpa

The llmsherpa project provides APIs to accelerate Large Language Model (LLM) projects. It includes features like LayoutPDFReader for PDF text parsing, smart chunking for vector search and Retrieval Augmented Generation, and table analysis. It is open-sourced under Apache 2.0 license.

2024-06-27 Tags: llm, pdf, text, parsing, retrieval augmented generation, foss, github, cpdomina by klotz

First / Previous / Next / Last / Page 1 of 0

SemanticScuttle - klotz.me

klotz: llm* + pdf*

Linked Tags

Related Tags