npm i llama-ocr
import { ocr } from "llama-ocr";
const markdown = await ocr({
filePath: "./trader-joes-receipt.jpg", // path to your image (soon PDF!)
apiKey: process.env.TOGETHER_API_KEY, // Together AI API key
});
We have a hosted demo at LlamaOCR.com where you can try it out!
This library uses the free Llama 3.2 endpoint from Together AI to parse images and return markdown. Paid endpoints for Llama 3.2 11B and Llama 3.2 90B are also available for faster performance and higher rate limits.
You can control this with the model
option which is set to Llama-3.2-90B-Vision
by default but can also accept free
or Llama-3.2-11B-Vision
.
This project was inspired by Zerox. Go check them out!