
Closed
Posted
Paid on delivery
I want to build an AI-driven knowledge-discovery platform focused on one rare disease. Essentially this would become a custom "medical intelligence" platform for a rare form of leukemia called T-PLL. This is not a diagnostic or treatment application and is not intended to provide medical advice. It is an AI-assisted research and surveillance platform whose purpose is to dramatically reduce the chance of missing important developments in research, clinical trials, or emerging therapies. Ideally, the platform would automatically search multiple public data sources on a scheduled basis, eliminate duplicate or low-value information, classify new findings by relevance (e.g., clinical trials, new publications, biotechnology developments, conference abstracts, investigator activity, etc.), maintain a searchable historical database, and generate weekly intelligence reports highlighting only meaningful changes. Core functions The platform has to 1) pull full-text research papers, conference abstracts, clinical-trial registry entries, biotech company announcements, major cancer-center press releases, and even relevant YouTube presentations, 2) extract structured data points (study design, cohort size, endpoints, biomarkers, funding source, etc.) with high accuracy, and 3) produce automated literature-review briefs that highlight trends and gaps. It should run these tasks on a schedule and push notifications when new material appears. The ideal candidate enjoys solving complex information management problems, can recommend the best technical approach rather than simply following instructions, and is interested in building a robust research assistant. I am not a researcher, a software engineer or even involved in the industry. This entire project is to help my dad who was recently diagnosed with this disease and I need more power in my searches. I am told by my own ChatGPT research that "this would involve a Python stack and leverage current NLP/LLM tooling (e.g., transformers, LangChain, GPT-4, spaCy), PDF parsing utilities, and APIs such as PubMed, CrossRef, [login to view URL], and YouTube Data API. A vector store (Pinecone, FAISS, or similar) for semantic search plus a lightweight web dashboard or Streamlit front end will keep things user-friendly." Deliverables • Crawling & ingestion pipeline covering all stated sources • Information-extraction module with clearly defined JSON/CSV schema • Automatic literature-review generator with citation linking • Real-time monitoring/alert service (email or Slack) • Web interface for search, filter, and export • Dockerised codebase with setup docs and a short demo video Acceptance criteria The system must capture at least 90 % of new items relevant to the disease and experimental or trial treatments, correctly extract the required data needed to track them, and generate a readable, reference-linked review on a regular basis (weekly or more frequently). Please outline your approach, and a rough timeline so we can move straight to milestones.
Project ID: 40613856
141 proposals
Remote project
Active 2 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
141 freelancers are bidding on average $1,134 USD for this job

Hi — Elias here from Miami. I see you're looking to build an AI-driven platform focused on rare disease research. This project has significant potential, but it comes with unique challenges. The real technical hurdle often lies in integrating diverse data sources and ensuring the AI models extract meaningful insights. What usually matters most is the quality and reliability of the data, along with the system's adaptability as new information emerges. The tricky part is managing the complexity of AI workflows while keeping the user experience intuitive. My approach would involve setting up a robust backend using Python and SQLite to handle data efficiently. I’d ensure the architecture is modular, allowing for future expansions and easy maintenance. This way, we can seamlessly integrate various APIs and AI models, including GPT-4 capabilities. I've worked on similar platforms where data extraction and analysis were critical, focusing on performance and reliability. A few questions to better understand the scope: Q1 – What specific data sources do you plan to integrate? Q2 – Are there particular user roles and permissions you envision for the platform? Q3 – How do you foresee the AI outputs being utilized or reviewed? Happy to discuss the details and suggest the best technical approach. Looking forward to hearing from you.
$1,200 USD in 6 days
8.2
8.2

Hello!, This is James from Hollywood... I read your project carefully, and I understand the goal: build an AI-driven knowledge discovery platform for one rare disease that can collect, organize, and surface useful research insights instead of just storing raw data. I’ve spent about 15 years working with Python, web apps, data extraction, NLP, API integration, and LLM-based systems, so this is the kind of project I’d take seriously from day one. My approach would be practical: 1) define the disease scope, data sources, and key questions 2) build the extraction and cleanup pipeline 3) structure the SQLite-backed knowledge base 4) add GPT-4 / LLM-assisted search, summaries, and insight generation 5) test with real research queries and refine the UX I’ve built similar research and data platforms where accuracy, traceability, and clean presentation mattered a lot, and I’d make sure this one feels reliable and useful. Could you please clarify the following questions to help me better understand the project? 1) Which rare disease should the platform focus on first, and do you already have preferred source websites, papers, or databases? 2) Do you want the AI to only summarize and connect findings, or also rank evidence and detect contradictions? 3) Should this be a private internal tool, or a polished public-facing web app with login and user roles? If helpful, I can also share a quick suggested architecture before we start so you can see exactly how I’d build it.
$1,200 USD in 5 days
6.9
6.9

Hi, I reviewed your request to build an AI-assisted knowledge-discovery platform for T-PLL that regularly ingests research, trials, biotech updates, and media, then turns new items into structured insights and weekly intelligence reports. I’ll implement the crawling & ingestion pipeline using Python with scheduled jobs and the relevant public APIs, then extract study design and biomarkers into a defined JSON/CSV schema. For knowledge discovery, I’ll integrate GPT-4 with LLM-based relevance classification and LLM Integration workflows, store embeddings in a vector store for semantic search, and generate citation-linked literature-review briefs. I’ll focus on high accuracy, deduplication, clean data pipelines, and reliable alerts so you never miss meaningful changes. Let’s discuss here now.
$750 USD in 30 days
6.6
6.6

I propose leveraging advanced NLP tools and APIs to create an AI-driven knowledge-discovery platform for T-PLL. Key features include data collection from PubMed, CrossRef, and others, semantic search using Pinecone or FAISS, and a user-friendly web interface. The development plan includes crawling & ingestion, information extraction, literature review generation, real-time monitoring, and a streamlined deployment process. The project is scheduled for completion within five months, with a commitment to meeting high-quality standards.
$1,350 USD in 5 days
6.5
6.5

As an AI and ML specialist, I understand the critical importance of accurate, efficient data processing. Having developed numerous systems like yours that balance vast amounts of information and present it in a easily digestible format, I can confidently state that my approach will ensure a 90% capture rate of relevant new items for your rare disease research platform. Using my comprehensive experience in Natural Language Processing and Python, we'll create a crawly and ingestion pipeline that pulls full-text research papers, conference abstracts, clinical-trial registry entries, biotech company announcements, major cancer-center press releases and even relevant YouTube presentations while eliminating duplicate or low-value information. What differentiates me? It's our capacity to merge AI with other facets seamlessly. My fluency in the Python stack grants me acuity in NLP/LLM tooling as well as PDF parsing utilities. This is crucial for extracting structured data points such as study designs, cohort size, endpoints, biomarkers, funding sources etc., with high precision.
$1,125 USD in 7 days
6.3
6.3

Hello! As per your project description, you are looking to build an AI powered research and intelligence platform dedicated to T PLL, designed to continuously monitor research papers, clinical trials, conference abstracts, biotech developments, cancer center updates, and relevant presentations. The platform would collect and organise information from multiple sources, identify meaningful new developments, remove duplicates and low value content, extract structured research data, and generate concise intelligence reports and alerts. My approach would focus on building a reliable research pipeline rather than simply connecting an LLM to search results. The system can automatically collect new information on a defined schedule, process documents and videos, classify findings by relevance, store historical knowledge, and use AI assisted extraction and summarisation to create searchable research insights. Given that this is intended as a research and surveillance tool rather than a diagnostic or treatment system, I would keep the platform clearly separated from medical decision making and include appropriate source references and confidence indicators throughout the research workflow. I would be glad to discuss your priorities and help define a practical first version that delivers useful intelligence without making the initial system unnecessarily complex. Best regards, Nikita Gupta
$1,125 USD in 7 days
5.6
5.6

Hi, I will deliver a Python-based AI platform for T-PLL research, crawling and extracting data from multiple public sources, with a functional web interface, I commit to completing this within the 750-1500 USD budget, can I start with a free sample? Waiting for your response in chat! Best Regards.
$1,125 USD in 3 days
5.3
5.3

? Hi , Thanks for sharing your project. I read through the description and it looks like you're looking for someone with experience in LLM Integration, Data Analysis, Python, Web Development, Data Extraction, SQLite, GPT-4, API Integration, Natural Language Processing and Research. This is the kind of work I do regularly, so I can focus on solving your problem instead of getting up to speed. I've worked on projects involving web applications, APIs, automation, and system integrations. Whether you're starting from scratch, improving an existing system, or fixing a specific issue, I'm comfortable jumping in wherever you need help. You can view examples of my work here: https://www.freelancer.com/u/thomasb726 Before we get started, I'd like to clarify a couple of details so we're on the same page and can avoid unnecessary back-and-forth later. I like to keep communication simple and transparent, so you'll always know where things stand. I'll keep you updated on progress, meet agreed milestones, and deliver a solution that's reliable, maintainable, and built with your long-term goals in mind. If it sounds like we're a good fit, I'd be glad to discuss the details and answer any questions you may have. Thanks for your time, and I look forward to hearing from you. Best, Tom
$1,000 USD in 3 days
5.3
5.3

I understand you need an AI-driven knowledge-discovery platform to track research, clinical trials, and emerging therapies for T-cell prolymphocytic leukemia (T-PLL). My experience building automated data aggregation and analysis pipelines for scientific literature has enabled clients to identify critical trends months ahead of competitors, a goal directly aligned with your need to reduce the chance of missing important developments. I will develop a Python-based platform leveraging libraries like BeautifulSoup and Scrapy for automated data scraping from PubMed, clinical trial registries, and relevant news sources. This data will be processed using NLP techniques with spaCy and stored in a PostgreSQL database. A user-friendly dashboard built with Streamlit will then visualize key findings, trends, and emerging research areas, providing scheduled summaries via email. How will the platform handle the ingestion of unstructured data from sources like research papers beyond simple abstracts? Ready to start as soon as you confirm scope.
$1,244 USD in 21 days
5.4
5.4

Hi there, I'm excited about the opportunity to work on the AI Platform for Rare Disease Research focused on T-PLL. Your vision for a comprehensive, AI-driven knowledge-discovery platform is both ambitious and vital, and I am eager to bring my expertise to this meaningful project. With a strong background in Python and natural language processing, I have successfully developed systems that integrate multiple data sources, including APIs like PubMed and ClinicalTrials.gov. My experience in data extraction and analysis ensures high accuracy in identifying and structuring key information from diverse inputs like research papers and clinical trials. I propose a modular approach where we first establish a robust data ingestion pipeline to gather relevant content. Using advanced NLP techniques and GPT-4, we'll develop an information-extraction module to parse and classify data effectively. A vector store will support semantic search capabilities, enabling efficient retrieval and analysis. Additionally, we'll create a user-friendly web interface using Streamlit, ensuring seamless interaction with the platform. My commitment is to deliver a solution that not only meets your requirements but also empowers you with actionable insights. I'm passionate about leveraging technology to drive impactful research, especially in areas as critical as rare diseases. Looking forward to collaborating with you. Best Regards,
$1,125 USD in 14 days
5.4
5.4

Hi, I understand this is not a general medical chatbot. You need a dependable research-surveillance system for T-PLL that continuously gathers new evidence, structures it, preserves source citations, and highlights only meaningful developments. I would build the first version in Python using official research APIs where available, robust PDF and text extraction, structured LLM outputs with validation, semantic search, and a simple web dashboard. Each finding would remain linked to its original source, with scheduled monitoring, duplicate detection, searchable history, and weekly intelligence reports. I have experience building AI, NLP, RAG, data-extraction, and research automation systems. Given the personal importance of this project, I would also keep the platform understandable and manageable for a non-technical user rather than delivering an overly complex research tool.
$1,350 USD in 35 days
5.1
5.1

With a strong background spanning over 14 years in Web and Mobile App Development, I'm confident in my ability to deliver exactly what you're looking for with this rare disease research project. I'm well-versed in programming languages like Python, PHP, and JavaScript-the essential tools required for this AI-driven knowledge-discovery platform. Through my career, I've undertaken projects ranging from educational solutions to online delivery platforms, and I see an opportunity within your project to truly make a difference. But beyond just my technical skills, what will set me apart is my empathy with your situation. Understandably, any ailment diagnosis can be trying for all family members involved. The fact that you've chosen to delve into the technical side with a drive to help your dad significantly resonates with me. Your project isn't solely about creating the platform; it's about extending support and hope to those combating T-PLL leukemia alongside. That emotional connection will manifest in the dedication and sensitivity invested into this development process - ensuring high-quality results that are user-friendly & SEO-friendly as mentioned.
$1,500 USD in 10 days
5.1
5.1

Hi, Aashiq (Ash) here from Cape Town, South Africa. This project instantly caught my eye, so I had to reach out. I see you’re looking to build an AI-driven platform specifically for T-PLL that automatically searches and organizes critical research and trial information. That’s a fantastic initiative to support your dad during this challenging time. I’ve helped numerous clients develop custom information systems that streamline research and enhance data accessibility. My approach focuses on leveraging Python, NLP, and robust data pipelines to ensure you receive timely, relevant updates. I’d be happy to share samples of similar projects that have successfully increased research efficiency. Based on what you mentioned, here is how we would approach the project: - Develop a robust crawling and ingestion pipeline for all specified sources. - Design an information-extraction module with a clear JSON/CSV schema. - Implement a literature-review generator for seamless reporting. - Set up a real-time alert service via email or Slack. - Create a user-friendly web interface for searching and exporting data. I’ll ensure clear communication throughout the process, delivering a seamless and user-focused solution optimized for performance. Best Regards, Aashiq
$1,350 USD in 14 days
4.9
4.9

Hi, I’ve carefully read your vision for a T-PLL medical intelligence platform, and this is exactly the kind of Python-driven research automation system I can help architect with confidence. I would build a robust pipeline for source ingestion, deduplication, relevance classification, structured Data Analysis, and scheduled reporting, then connect it to a practical Web Development dashboard for search, filtering, alerts, and exports. My experience aligns well with custom platforms that combine Python, APIs, workflow automation, and user-friendly interfaces. For your use case, I’d recommend a modular approach: data collectors for PubMed/ClinicalTrials/YouTube and other public sources, extraction and normalization into a clean schema, semantic ranking for important updates, and a reporting layer that generates weekly intelligence briefs with citation links. I can also deliver the Dockerised setup, documentation, and demo flow you requested. I can outline milestones immediately and begin with architecture plus source-mapping in the first few days. Which source should be treated as the highest-priority signal first so I can structure milestones around maximum early value? Best regards, KANIKA
$1,500 USD in 30 days
4.5
4.5

Hi, I saw your project and think I can deliver what you need. Let's build a system so smart that your future self sends us both a thank-you note. I have read your project description. I am the one who delivers fast and quality work, and communicates until the result feels right. I am a reliable freelancer with 10 years of experience in Python, Web Development, API Integration, SQLite. I have helped many clients achieve great outcomes and finish projects smoothly. Visit my profile to check my latest work and read short client reviews. Please reach out in chat so we can align on your needs and goals. Looking forward, Yuan
$1,000 USD in 4 days
4.2
4.2

Building an AI-driven platform for T-PLL is crucial for tracking breakthroughs in research. Your goal of automating data retrieval and analysis can drastically enhance your ability to stay informed. I will create a robust system using Python and the latest NLP techniques to efficiently monitor diverse sources. You’ll benefit from: - Timely alerts on significant developments, reducing the chance of missing vital information. - Structured insights that pinpoint trends and gaps, saving you hours of manual searching. **Here’s how we’d approach it:** - Develop a comprehensive crawling pipeline for all specified sources. - Implement high-accuracy data extraction modules. - Create an automated literature review generator with user-friendly outputs. We're newer to Freelancer but bring 9+ years of experience delivering similar projects off-platform. I’d love to offer a no-obligation consultation to discuss your specific needs further. What’s your timeline for launching this platform? Kind regards, Trichelle
$900 USD in 7 days
3.4
3.4

Hello! Welcome to NexoraTec! I understand your need for an AI-driven knowledge-discovery platform focused on T-PLL, a rare form of leukemia. Our team specializes in solving complex information management problems and recommending the best technical approach for building robust research assistants. We are committed to helping you create a custom "medical intelligence" platform that streamlines research, clinical trials, and emerging therapies related to T-PLL. Our approach involves leveraging a Python stack and cutting-edge NLP/LLM tools like transformers and GPT-4, along with APIs such as PubMed and ClinicalTrials.gov. We will develop a crawling & ingestion pipeline, information-extraction module, literature-review generator, real-time monitoring/alert service, web interface, and Dockerised codebase to meet your requirements. We are dedicated to ensuring that the system captures 90% of relevant new items, extracts necessary data accurately, and generates readable reviews regularly. Let's discuss further details and create a timeline to achieve your project milestones. Happy to plan the best way with you. Thanks!
$807 USD in 8 days
3.1
3.1

Hello there. I hope you are donig well. I have successfully developed complex AI-driven platforms that streamline research processes, including web scraping for academic resources and automated data extraction. My background in Python and data management equips me with the skills necessary to tackle this project effectively. I understand that you need an AI-assisted research platform for T-PLL that aggregates and analyzes information from diverse sources. I will create a robust system that ensures relevant data is captured, organized, and reported, minimizing the risk of missing important developments in research and therapies. I will provide a high-quality platform with a user-friendly web interface, a comprehensive crawling and ingestion pipeline, and an automated literature review generator. Utilizing the best technologies, I will ensure the solution is efficient, scalable, and tailored to your specific needs, delivering a product that meets your expectations. Please feel free to reach out to me. I look forward to working with you. Best regards, Billy Bryan
$1,050 USD in 7 days
3.1
3.1

Hello, "RAG-Based Medical Knowledge Engine" — you need a system to synthesize T-PLL research papers into actionable insights. I will implement a Retrieval-Augmented Generation (RAG) pipeline using LangChain and a vector database to ensure the AI only cites verified medical literature. This prevents hallucinations and provides traceable sources for every claim. I previously built an automated AI-driven video backup system here: https://www.freelancer.com/projects/automation/Automate-Reolink-Video-Backup-Script/reviews Will the primary data sources be provided as a set of PDFs, or should the platform scrape PubMed and other medical databases in real-time? Looking forward to working with you. Artur Giżycki
$1,200 USD in 14 days
3.2
3.2

As an experienced full-stack developer with a particular focus on AI-driven systems and data management, I have the skills and knowledge necessary to create your AI platform for rare disease research. I understand the crucial importance of thoroughly parsing large amounts of potentially relevant information, and have built similar information-retrieval systems in previous projects. In line with your desired technology stack, I am well-versed in using Python - from working with various NLP/LLM tooling like transformers to understanding PDF parsing utilities. This project aligns completely with my commitment to delivering clean, production-ready work on deadline even under substantial pressure like that of an Enterprise Fintech scenario. I guarantee you zero late deliveries. My last 5 years of development experience combined with my understanding of OpenAI Agents SDK, LangChain, and RAG pipelines will ensure I construct a high-functioning solution that precisely meets your objective: capturing at least 90% of new findings, then correctly extracting and documenting your research before producing a reference-linked review, all on a routine basis. Treasure your dad’s wellbeing by hiring someone who will put his utmost effort into building this platform; choose me – ABUZAR
$750 USD in 7 days
3.1
3.1

Belton, United States
Member since Jun 5, 2024
₹12500-37500 INR
$250-750 USD
$250-750 USD
$30-250 USD
$15-25 USD / hour
$250-750 USD
€20000-50000 EUR
₹1500-12500 INR
$750-1500 USD
₹750-1250 INR / hour
$60-80 AUD
₹1500-12500 INR
$250-750 USD
$250-750 USD
$30-250 USD
$30-250 USD
₹1500-12500 INR
₹600-1500 INR
$2-8 USD / hour
₹750-1250 INR / hour