
Closed
Posted
Paid on delivery
I am looking for an experienced Transkribus / Handwritten Text Recognition (HTR) specialist to help develop a high-quality transcription model for a large collection of handwritten family diaries. The collection consists of several thousand pages of vineyard and farming diaries from South Australia, spanning approximately 1890–1947. The diaries record daily vineyard work, livestock management, crop planting, weather, visitors, grape varieties, local events and family history. The ultimate goal is to create a reliable transcription model that can be used across the entire collection. Rather than transcribing every page manually, I am seeking someone who can: - Review a representative sample of diaries. - Determine an appropriate training and validation approach. - Train a Transkribus (or equivalent HTR) model. - Demonstrate that the model can accurately transcribe previously unseen diary pages. I am interested in working with someone who understands both: - historical handwriting and transcription, and - training and evaluation of handwritten text recognition models. I can provide a representative subset of diaries covering different periods and handwriting styles. I expect to retain a separate set of diaries and pages that will not be used during training so that the model can be tested independently. I am open to your recommended approach, but I would expect something along the lines of: - A trained HTR model. - Documentation of the training process. - Validation results against pages not used during training. - Recommendations for further improving the model. - A demonstration of transcription quality on previously unseen diary pages. This will likely be a two-phase engagement: (1) pilot assessment and proof of concept, followed by (2) full implementation if the pilot is successful. Please tell me: 1. Your experience with Transkribus or other HTR platforms. 2. Any previous historical manuscript projects you have worked on. 3. How you would approach training and validating a model for a large collection like this. 4. How you would measure success. 5. Examples of similar projects, if available.
Project ID: 40576149
42 proposals
Remote project
Active 4 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
42 freelancers are bidding on average $2,102 AUD for this job

Hi, I'd be happy to help with your HTR project. I have experience working with AI-powered document processing, OCR/HTR workflows, historical document digitization, and model evaluation. My approach would begin with a pilot phase to assess handwriting variations, prepare high-quality ground truth, train a Transkribus (or equivalent) model, and validate it against a separate set of unseen diary pages to measure real-world accuracy. You'll receive a trained model, documented training workflow, validation results, recommendations for improving accuracy, and a clear success metric based on character and word error rates. If the pilot meets the target accuracy, we can confidently scale the model across the full diary collection. I look forward to discussing your project. Best, Muhammad Usman
$1,850 AUD in 3 days
6.6
6.6

⭐⭐⭐⭐⭐ Create a Reliable HTR Model for Handwritten Family Diaries ❇️ Hi My Friend, I hope you are doing well. I've reviewed your project requirements and noticed you're looking for an experienced Transkribus/HTR specialist. Look no further; Zohaib is here to help you! My team has successfully completed 50+ similar projects in handwritten text recognition. I will review your diaries, determine the best training methods, and create a reliable transcription model that meets your needs. ➡️ Why Me? I can easily develop your transcription model as I have 5 years of experience in HTR and historical handwriting analysis. My expertise includes data collection, model training, and validation processes. Additionally, I have a strong grip on data analysis, ensuring your model performs well across various handwriting styles. ➡️ Let's have a quick chat to discuss your project in detail and let me show you samples of my previous work. Looking forward to discussing this with you in our chat. ➡️ Skills & Experience: ✅ Transkribus ✅ Handwritten Text Recognition ✅ Historical Manuscript Analysis ✅ Data Collection ✅ Model Training ✅ Validation Techniques ✅ Data Analysis ✅ Documentation ✅ Transcription Quality Assessment ✅ Project Management ✅ Quality Improvement Strategies ✅ Communication Skills Waiting for your response! Best Regards, Zohaib
$1,800 AUD in 2 days
5.6
5.6

Hi there, I can help develop a robust HTR workflow for your historical diary collection by first assessing a representative sample to identify handwriting variations, document quality, and transcription challenges before defining an effective training and validation strategy. My approach would include preparing and normalizing the training data, creating accurate ground-truth transcriptions, training and iteratively refining a Transkribus (or equivalent HTR) model, and evaluating its performance using a completely independent test set to ensure realistic accuracy on previously unseen pages. I will document the complete process, report recognition metrics such as Character Error Rate (CER) and Word Error Rate (WER), identify recurring transcription issues, and provide clear recommendations for further improving model performance as the project scales to the full collection. Could you also confirm whether you already have any manually transcribed pages available for training, or should the initial ground-truth dataset be created as part of the pilot? Regards, Bryson.
$2,500 AUD in 14 days
5.5
5.5

I can develop a high‑accuracy HTR model for your historical vineyard diaries using Transkribus or an equivalent modern handwriting recognition pipeline. I have experience working with heterogeneous handwriting, archival material and custom model training. If helpful, I can explain how I design an HTR training pipeline or how I run historical validation on unseen pages. My approach: • Review a representative subset to map handwriting styles, ink variations and structural patterns across decades. • Define a strict training/validation split to ensure real generalization (holdout diaries never used in training). • Train a custom Transkribus model optimized for 1890–1947 handwriting. • Evaluate performance on unseen pages with clear metrics (CER/WER, stability across periods, sensitivity to style shifts). • Provide documentation of the full training process, validation results and recommendations for further improvement. • Demonstrate transcription quality on pages not included in the dataset.
$2,000 AUD in 7 days
5.0
5.0

Hello Sir/MAM I am a skilled full stack developer. Having rich experience in Java , C++ , C , C# , Python , Eclipse , Sql , Mysql , .Net ,Oracle , Object Oriented Programming , Data Structure , Algorithms, Linux , Windows , Cloud , Azure . I have a perfect grip on “Artificial Intelligence” “Automation” , and work in “Machine Learning” Deep Learning “Computer Vision ”. My track record as demonstrated in my 100% job completion and 5-star review rating showcases My ability to deliver exceptional results on time and with utmost quality I believe that my skill set makes me the ideal candidate for this project Please come on chat we will discuss more about this I will be waiting for your reply . Thanks and Best Regards
$1,551 AUD in 4 days
4.2
4.2

What'll decide this is whether one model can span 1890–1947, or whether the handwriting shifts enough across writers and decades that you need a few. Several thousand pages over ~57 years usually means more than one hand, and Transkribus CER swings hard between a neat 1890s clerk and a shakier later one. So on the pilot I'd sort the sample by writer and period first, then transcribe enough ground truth per hand to see where character error rate lands. Grape varieties, place names and farm terms won't exist in any base language model, so I'd build a domain dictionary from the transcribed text. On the held-out pages the split matters — for an honest CER the test set needs the writers the model trained on, plus a couple it's never seen to show its real ceiling. Are the diaries mostly one hand, or several across the years? That decides whether the pilot proves one model or a small set.
$1,999 AUD in 20 days
2.5
2.5

❤️Hi there❤️ Your project is a perfect match for my expertise in Transkribus and HTR platforms. I have successfully worked on historical manuscript projects in the past, blending transcription and model training seamlessly. My approach involves meticulously reviewing the diaries, devising a robust training plan, and validating the model for accuracy on unseen pages. I measure success by delivering a trained HTR model, detailed documentation, validation results, and recommendations for enhancements, ensuring top-notch transcription quality. This project perfectly aligns with my skills, and I am excited to collaborate to bring your vision to life. Please take a moment to review my profile for more insights. I look forward to discussing this further with you. Warm regards, Thaveesha
$1,500 AUD in 7 days
1.4
1.4

Hello! I’ve worked on a project that involved building a handwritten text recognition model for historical documents, leading to a 90% accuracy rate in transcription. I can share the details and results in chat. For your project, I’d start by reviewing a sample of the diaries to understand the handwriting styles, then determine a training strategy using Transkribus. I’d ensure that we set aside a validation set to properly test the model’s performance on unseen pages, which is crucial for reliability. What specific handwriting challenges do you anticipate with the different styles in the diaries? If you’re open, I can share my similar build and we can see if it fits your needs.
$2,250 AUD in 7 days
1.0
1.0

Hi there, I’ll build an HTR transcription model for your 1890-1947 vineyard diaries using Transkribus (or an equivalent stack) and a training/validation setup that actually tests unseen handwriting. That auth part might be annoying here. The risky part is making sure the model generalizes across handwriting styles, so we’ll use a held-out diary subset per period and measure accuracy on truly unseen pages. I’ll handle this in two phases: sample review + labeling plan, then training, tuning, and evaluation. - Review the provided sample pages and decide layout/segmentation strategy - Train the model and run validation on withheld pages - Produce a transcription demo on new pages plus a short training report I have hands-on experience with HTR workflows and historical text cleanup. I’m available to start immediately and can share pilot results quickly. Slavko
$1,500 AUD in 6 days
0.0
0.0

Re: Handwritten Text Recognition Model for Historical Diaries My initial assessment for creating a high-accuracy HTR model for your 1890–1947 diaries suggests a stack using Python with TensorFlow and AWS for scalable model training and evaluation. I have direct experience implementing custom recognition models for historical documents. For instance, on a past project, I successfully developed and trained a custom NLP model to digitize and classify a large corpus of cursive legal documents. The model achieved a Character Error Rate (CER) of under 4% after fine-tuning on a specific scribe's handwriting, significantly reducing manual transcription efforts. I propose the following key steps: 1. Analyze a representative sample of diary pages to identify handwriting styles and ink degradation to inform the data augmentation strategy. 2. Establish a baseline model using a pre-trained HTR engine, then fine-tune it with a manually transcribed training set from your provided diaries. Happy to elaborate on my approach. Regards, Anton K.
$1,500 AUD in 7 days
0.0
0.0

Hi, I reviewed your project: Build a Handwritten Text Recognition Model for a Historical Vineyard Diary Collection. I can help you build a practical AI-powered solution with secure API integration, clean backend architecture, automation workflows, database design, and a production-ready admin/dashboard system. My experience includes AI assistants, OpenAI/LLM integrations, RAG/knowledge-base workflows, Laravel/PHP, React, Node.js, APIs, databases, and deployment. Please message me so I can confirm the workflow, data sources, integrations, and success criteria before we start. Portfolio: https://www.freelancer.com/u/irfanui Regards, Mohammad 4th Dimension Partners
$2,200 AUD in 20 days
0.0
0.0

Hi, I've read through your project details, and it sounds like you need a reliable Handwritten Text Recognition model to transcribe a rich collection of historical vineyard diaries from South Australia. My approach would begin with reviewing a sample of the diaries to understand the various handwriting styles. From there, I'd set up a training and validation plan, leveraging my experience with Transkribus and similar HTR platforms. With 7+ years in this field, I've worked on projects involving historical manuscripts, where understanding the nuances of handwriting is crucial. I can ensure that the model I develop accurately transcribes unseen pages, providing documentation and validation results as you requested. One question I have is: are there specific handwriting styles or periods within the diaries that you think will be particularly challenging for the model?
$1,500 AUD in 21 days
0.0
0.0

Hi, I am a software engineer with over 16 years of experience, including OCR/HTR workflows, document AI, model training/evaluation and production-quality data pipelines for messy real-world records. For this diary collection I would treat the first phase as a careful pilot: inspect pages across the 1890-1947 range, group handwriting/style changes, define train/validation/held-out sets, then train and tune a Transkribus or equivalent HTR model and report CER/WER on pages never used in training. I have worked with historical and low-quality handwritten/printed archives where the main challenge was not just model training, but choosing representative ground truth and avoiding optimistic validation. My approach would include documenting the sample selection, baseline accuracy, training iterations, error patterns, and practical recommendations for scaling to the full vineyard diary collection. Success should be measured by character/word error rate on the independent test pages, plus readable output quality for names, dates, vineyard terms, weather entries and recurring farming vocabulary. Similar archive/document projects cannot be shared publicly, but I can show/discuss relevant private examples if needed. A few questions: how many manually transcribed pages already exist, and are the scans already segmented/exported for Transkribus? Please contact me to discuss details.
$2,800 AUD in 21 days
0.0
0.0

Hello, I can build and validate a Transkribus HTR model tailored to your South Australian vineyard diaries from 1890 to 1947. I trained a Transkribus model on 1,800 pages of 19th century diaries and achieved 92 percent character accuracy on a withheld test set. I will review your representative sample, create accurate line level ground truth, split data into training, validation and a held out test set you retain, train and iterate models in Transkribus, produce documented preprocessing and training notes, and deliver validation metrics plus sample transcriptions of unseen pages. Success will be measured by character error rate, word accuracy and qualitative recovery of domain terms such as grape varieties and dates. Can you provide a representative sample of 200 to 300 pages and confirm a withheld test set for validation? Happy to jump on a quick chat. Ali Zain
$2,250 AUD in 7 days
0.0
0.0

⚡ I can help develop a reliable handwritten text recognition model that transforms your historical diary collection into searchable and structured digital records. Hello, I have experience with AI model development, OCR/HTR workflows, data preparation, machine learning evaluation, and building practical AI solutions for complex document processing tasks. I can analyze a representative sample of your vineyard diaries, evaluate handwriting variations, prepare training and validation datasets, and develop a Transkribus-based HTR model optimized for your historical manuscripts. My approach would include iterative training, independent testing on unseen pages, accuracy measurement using recognized metrics, and documentation of the complete workflow for future expansion. I understand the importance of preserving historical context and would focus on achieving reliable transcription quality while reducing manual effort across thousands of pages. I would be happy to discuss the pilot phase, evaluation criteria, and a scalable roadmap for the full collection. Best regards, Darial C.
$1,500 AUD in 7 days
0.0
0.0

Howdy! I have worked extensively with handwritten text recognition pipelines, including Transkribus and custom HTR workflows built on top of PyTorch and TensorFlow with CRNN and Transformer based architectures. While my primary background is in AI engineering rather than paleography, I have applied these tools to document digitization tasks involving degraded, historical, and non standard handwriting, including log books, ledgers, and archival records where layout segmentation and character level accuracy were critical. I understand the full pipeline from page preprocessing and baseline detection through ground truth preparation, model fine tuning, and CER/WER evaluation against held out validation sets. Thank you. Marcos.
$2,250 AUD in 7 days
0.0
0.0

Adelaide, Australia
Member since Jul 11, 2026
$15-25 USD / hour
₹600-1500 INR
$3000-5000 USD
₹1250-2500 INR / hour
$5000-10000 USD
₹15000-20000 INR
$15-25 USD / hour
$2-8 USD / hour
₹1500-12500 INR
₹750-1250 INR / hour
$8-15 USD / hour
$10-30 USD
₹600-1500 INR
$30-250 USD
₹600-1500 INR
₹1500-12500 INR
₹1500-12500 INR
₹10000-200000 INR
$250-750 AUD
$1500-3000 AUD