Employee access required
This tool is marked employee-only. Sign in with an employee account from the main Omni-Science site to access it.
Time's Up!
Your free time for this tool has been used. Sign in for more time or upgrade to Omni-Pass for unlimited access.
Get Omni-PassOmni-Convert To Text

Getting ready for greatness
Initializing Systems...

Owned by Omni-Science
Drop Files or ZIPs here
Supports: Text formats, Code, PDF, Office, Images, Audio, Video, ZIPs.
Selected Files:
Crawler
Depth 1: Scrapes Start Page + Direct Links.
Infinite: Scrapes everything found (Same Domain).
Crawler is locked to the original domain to prevent internet-wide loops.
Requires permission to scan available devices.
Click to Start Dictation
00:00
Dictation is now processed by Omni-Bot through the secure Cloudflare backend. No local AI model is downloaded.
As you speak and pause, audio chunks are safely transcribed locally. No data is sent to external servers.
Drag & Drop files or ZIPs. Supports mixed formats (e.g., ZIP with PDFs & Code).
- Enable OCR checkbox to extract text from images.
- Enable speech to text to natively extract audio from MP4/MP3s and transcribe it.
Enter a URL to scrape text.
- Crawler Mode (Toggle): "ON" follows links on the page. "OFF" scrapes only the specific URL provided.
- Depth Slider: Determines how many clicks away from the start page the crawler will go. "1" = Start Page + Direct Links.
- Multi-URL: Use the (+) button to batch process multiple different websites at once.
Public: Paste URL. No token needed (uses Proxy).
Private: Paste URL + Token (uses API).
Use the microphone under the "Voice" tab to talk directly to the application. It will record your voice, send it to the AI for processing, and convert the returned transcript into a formatted text file.
All processed content is merged into a single .txt file with a clean directory tree at the top.
Code & Config: Any text-based format including custom, made-up, and unknown extensions (e.g., .njk, .custom, no extension).
Documents: PDF, DOCX (Word), PPTX (PowerPoint), XLSX (Excel).
Images (OCR): JPG, PNG, BMP, WEBP (Requires "Enable OCR").
Audio/Video (STT): MP3, MP4, WAV, OGG, WEBM, M4A, etc. (Requires "Enable speech to text for audio and video files". Processed locally via WASM).
Web: Single page or Full Site Crawling (Depth control).
Archives: ZIP (Recursively processes all above formats inside).
100% Serverless Processing.
This entire application runs without a custom backend server.
- PDF parsing via pdf.js
- OCR via Tesseract.js (WASM)
- Speech-to-Text (A/V) processed through Omni-Bot on the secure Cloudflare backend
- Web fetching via Omni-Science Cloudflare edge
- Zero backend data retention