[verdict][prompt curatoriat][sursă: scantoexcel.ai][traducere: la coadă]
Pot să vibecodez ScanToExcel?
// se construiește într-un weekend, dar rămân goluri reale
A vision model will read a clean printed table on the first try, and that demo is genuinely an afternoon of work. Consistency is the part that is not. Real documents arrive with merged cells, multi-line rows, columns that shift between pages and numbers that must survive as numbers, and a one-shot prompt handles each of those differently every time you run it. ScanToExcel puts a purpose-built extraction pipeline between the model and the spreadsheet precisely because the model alone is not reproducible. You can copy the easy half of this product in a sitting and spend months on the half that makes it trustworthy.
încredere: medie
ce e: Photograph a table or form and get a real .xlsx back, handwriting included
Indice de constructibilitate · joc editorial
- Preț 39,99 $/lună greutate: plus 3
- Timp o sesiune pentru tabele curate; timp nedefinit pentru consistență greutate: plus 3
- Categorie documente și pdf nu cântărește
- Moat modele proprietare · calitatea execuției greutate: minus 3
- Încredere medie greutate: zero
- Ce pierzi 5 elemente greutate: minus 2
- Site scantoexcel.ai (si apre in una nuova scheda) nu cântărește
Se anulează. E cazul în care decide cât de tare te enervează abonamentul.
Gioco editoriale: il verdetto dice se un agente può, l’indice se conviene.
Ce construiești
Upload a photo or PDF page, send it to a vision model asking for the table as structured rows, and write the result to an .xlsx file.
ce îți trebuie
- OpenAI or Anthropic API key with vision support, in .env
- Node.js 22
- A spreadsheet writer such as SheetJS or ExcelJS
Promptul
Un weekend cu un agent de coding. Golurile care rămân sunt mai jos, la „ce pierzi”.
Build a document-to-spreadsheet converter inspired by ScanToExcel.
Use exactly this stack: Next.js 15 + TypeScript.
Primary job: the user uploads a photo or a PDF page containing a table, and gets back a downloadable .xlsx whose cells match the document.
Start from an empty folder and create the complete working project.
Send the image to one vision-capable model and ask for the table as structured rows and columns, not as prose.
Write the result with a spreadsheet library so numbers arrive as numbers and dates as dates, not as text.
Show the extracted table in the browser for review and let the user fix cells before downloading - never hand over a file the user has not seen.
Handle multi-page PDFs by processing pages in order and dropping a repeated header row when the same header appears on a later page.
Single-user and private by default; delete uploads once the download is produced, and say so in the UI.
Put every secret in .env and provide .env.example.
Include clear empty, loading, validation, success and failure states, and show the model's confidence when it reports one.
Accessible keyboard navigation, labels, focus states and sensible contrast.
Deliberately exclude these paid-product advantages: a native mobile capture app, tuned handwriting recognition, batch processing of very long documents.
State plainly in the README that handwriting and low-quality photos are where this build degrades.
Write unit tests for the row-parsing logic and one end-to-end smoke test that converts a sample image to a valid .xlsx.
Create a README with setup, architecture, per-page cost estimate and limitations.
Run the tests and build before finishing, then fix what fails. Prompt curatoriat: scris și revizuit manual pentru această aplicație. În engleză intenționat — e limba în care agenții de coding se descurcă cel mai bine.
Ce pierzi
- a purpose-built extraction pipeline rather than raw model output
- the same document producing the same spreadsheet twice
- structure held across pages: merged cells, multi-line rows, shifting columns
- reliable handwriting recognition
- a phone app that captures and converts without a laptop
De ce se plătește în continuare
Because the demo works and the hundredth document does not. Accountants feed it crooked phone photos of carbon-copy forms with merged headers, and the difference between a tool and a script is what happens on that page. Paying for output you can rely on without checking every cell is an easy trade for someone billing hourly.
moat: Modele proprietare Calitatea execuției ce e un moat
custom extraction pipeline, reproducible structure handling
Cine l-a construit deja
Să pornești de aici e tot vibecoding: promptul e pentru când o vrei exact în felul tău.
- img2table (se deschide într-o filă nouă) — Table identification and extraction from images and PDFs, no model API required.
- Camelot (se deschide într-o filă nouă) — Extracts tables from text-based PDFs into DataFrames.
- docling (se deschide într-o filă nouă) — Document parsing to structured formats, including table structure recovery.
Ești de acord?
Balanța voturilor
Ancora nessun voto: il tuo è il primo.
Nessun voto ancora — il primo pesa.