
Indexify

Click to visit website
About
Indexify is an open-source data framework designed for effortless ingestion and extraction of unstructured data at any scale for LLMs. It features a real-time extraction engine, pre-built extractors for various data types (documents, presentations, videos, audio), and supports custom extractor creation. Data retrieval is facilitated by semantic search and SQL querying. Indexify scales from local runtimes to large-scale Kubernetes deployments across multiple clouds. It also provides end-to-end observability and monitoring of ingestion, extraction, and retrieval processes.
Platform
Task
Features
• semantic search
• multi-modal support
• runs on laptops and across large-scale deployments (kubernetes, vms, bare metal)
• sql querying
• custom extractor creation using sdk
• reliable extraction for unstructured data (documents, presentations, videos, audio)
• pre-built extraction adapters
• real-time extraction engine
Job Opportunities
Founding Applied AI Scientist
Indexify is an open-source, real-time data extraction framework for LLMs, supporting various data types and scalable deployments.
Benefits:
401(k) plans
Comprehensive Healthcare and Dental Benefits
Education Requirements:
Ph.D. or Bachelor's degree in a quantitative field such as Computer Science, Mathematics, or equivalent industry experience
Experience Requirements:
4+ years of experience working with AI/ML models, specifically in the fields of document understanding, computer vision, and multi-modal learning
Proven expertise in training and evaluating models for complex document extraction
Deep NLP Expertise
OCR Integration
Model Pretraining and Fine-tuning
Other Requirements:
Solid programming skills in Python and proficiency in at least one deep learning framework (e.g., TensorFlow, PyTorch)
Layout Analysis
Benchmarking and Evaluation
Vision-Language Models
Responsibilities:
Design, train, and evaluate document understanding models for extracting complex data
Develop and optimize multi-modal visual Q&A models
Collaborate with the team to integrate AI-driven features into Tensorlake’s platform
Work closely with users and customers to understand their needs
Show more details
Founding Backend Engineer
Indexify is an open-source, real-time data extraction framework for LLMs, supporting various data types and scalable deployments.
Benefits:
401(k) plans
Comprehensive Healthcare and Dental Benefits
Education Requirements:
Ph.D. or Bachelor's degree in Math, Computer Science, or other quantitative fields, OR equivalent experience
Experience Requirements:
7+ years of relevant work experience
Experience in building large-scale distributed systems
Other Requirements:
Knowledge of systems programming languages such as Rust, Go, C++, or C
Designing observable systems that operate at internet scale
Deep knowledge of operating and using cluster schedulers
Responsibilities:
Design and implement a distributed control plane for operating Indexify on public clouds
Design and implement workflows for cluster operations and bootstrapping in VPCs
Focus on long term operability of the system and services
Work closely with the Founder on the company's technical direction and platform
Work closely with our users to learn the impact of our product and improve their experience
Show more details
Founding Product Engineer
Indexify is an open-source, real-time data extraction framework for LLMs, supporting various data types and scalable deployments.
Benefits:
Healthcare, Dental and Vision Insurance
401(k) plans
5 weeks of PTO
Experience Requirements:
At least 7 years of front-end or full-stack development
Familiarity with technologies such as Python, React, Typescript, FastAPI, or SQLAlchemy
Other Requirements:
Motivated people who are excited to build tools to power the next generation of cloud applications
Passionate about working adjacent to users and the product
Responsibilities:
Develop delightful UIs or high quality backend business logic that empower software developers and simplify programming
Work with a team of leading distributed systems and machine learning experts
Communicate your work to a broader audience through talks, tutorials, and blog posts
Help us to build and shape a world class company
Show more details
Ratings & Reviews
No ratings available yet. Be the first to rate this tool!
Alternatives

LiftData
LiftData provides real-time AI-powered data extraction from various content sources using a decentralized, scalable platform.
View Details
Lido
Lido is an AI OCR tool that converts PDFs to Excel, accurately extracting data from any PDF or email into a spreadsheet. Automate manual data entry and reduce errors with this #1 AI OCR tool.
View Details
SynerAI
SynerAI uses advanced NLP and generative AI to extract data and insights from the world's news, providing valuable datasets and market insights to financial institutions and consultants.
View Details
Gilio
Gilio processes documents with AI, extracting and transforming information for automation. It integrates with various systems via API and offers features such as data validation, document digitization, and workflow automation.
View Details
ASSIST
ASSIST is a document management software that automates data entry and streamlines AP & AR categorization. It keeps your financial records in order with easy extraction and reporting.
View DetailsFeatured Tools
Songmeaning
Songmeaning uses AI to reveal the stories and meanings behind song lyrics. It offers lyric translation and AI music generation.
View DetailsWhisper Notes
Offline AI speech-to-text transcription app using Whisper AI. Supports 80+ languages, audio file import, and offers lifetime access with a one-time purchase. Available for iOS and macOS.
View DetailsGitGab
Connects Github repos and local files to AI models (ChatGPT, Claude, Gemini) for coding tasks like implementing features, finding bugs, writing docs, and optimization.
View Details
nuptials.ai
nuptials.ai is an AI wedding planning partner, offering timeline planning, budget optimization, vendor matching, and a 24/7 planning assistant to help plan your perfect day.
View DetailsMake-A-Craft
Make-A-Craft helps you discover craft ideas tailored to your child's age and interests, using materials you already have at home.
View Details
Pixelfox AI
Free online AI photo editor with comprehensive tools for image, face/body, and text. Features include background/object removal, upscaling, face swap, and AI image generation. No sign-up needed, unlimited use for free, fast results.
View Details
Smart Cookie Trivia
Smart Cookie Trivia is a platform offering a wide variety of trivia questions across numerous categories to help users play trivia, explore different topics, and expand their knowledge.
View Details