
In Progress
Posted
Paid on delivery
I need a comprehensive, research-ready weather dataset that covers the entire People’s Republic of China—mainland provinces as well as Hong Kong and Macau—from 1 January 2008 right up to the day you hand the files over. Core content • Mandatory variables: daily minimum and maximum temperatures for every available station or grid point. • Additional variables: please include precipitation totals, relative humidity, and wind speeds, plus any other parameters that a publicly accessible source provides (sunshine hours, pressure, visibility, etc.). The more complete the record, the better. Format & structure Everything should arrive in tidy CSV files. A single national file is fine if it remains manageable; otherwise split logically (e.g., by province or year) but keep a consistent column schema. I will also need: 1. A concise README explaining data sources, retrieval methods (API, scraping, bulk download, etc.), time zones, and any unit conversions you performed. 2. A data dictionary defining each column and its units. 3. A short script (Python or R is perfect) that reproduces the data pull or updates it in the future. Quality expectations • Dates must be continuous with no silent gaps—flag any missing days explicitly. • Use numeric types where appropriate; don’t mix units. • Station metadata (lat, lon, elevation, name, province) should accompany the measurements or be supplied in a separate file. • Only publicly licensed or otherwise shareable sources, suitable for academic research, are acceptable. Delivery timeline is flexible within reason, but let me know early if certain regions or periods prove difficult so we can decide on substitutions or workarounds. If you have existing archives that already satisfy these criteria, feel free to propose using them.
Project ID: 40639512
25 proposals
Remote project
Active 6 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
25 freelancers are bidding on average $106 USD for this job

Hello, I have thoroughly reviewed the project requirements for compiling a comprehensive weather dataset for China from 2008 to present, including daily temperature, precipitation, humidity, and wind speed data. Let's chat and discuss it further. To handle your project, I will start with sourcing data from publicly accessible databases and APIs, ensuring the inclusion of mandatory variables like temperature, while also collecting additional parameters such as precipitation and humidity. I will organize the data into tidy CSV files, with a clear README, data dictionary, and a script for future updates. The deliverables will include a complete weather dataset in CSV format, a detailed README file, a data dictionary, and a script for data retrieval. Before signing-off my bid, I would like to ask a question, i.e., have you identified specific sources for the weather data or would you like me to select them based on my expertise? Best Regards, Aneesa.
$100 USD in 1 day
7.1
7.1

Hi, Data pulls like this live or die on the source plan, so let me flag the main tradeoff early. CMA station data before roughly 2015 has real gaps and licensing limits, while gridded reanalysis like ERA5 gives continuous 2008 to now coverage for temperature, precip, humidity, wind, and pressure with clean lat/lon metadata. For research use I'd usually base the backbone on a shareable reanalysis source and layer public station records where they exist, flagging every missing day explicitly rather than silently filling. I write Python daily and handle scraping, API pulls, and data processing. Quick question: do you want true station observations where available, or is a consistent gridded product acceptable for the whole period? I'd start on a milestone so you only release once the first province file and README check out. Adil
$114.40 USD in 7 days
6.0
6.0

Hi Angela, I will deliver a comprehensive China weather dataset from 2008 to present in CSV files with mandatory variables and additional parameters. I commit to a reasonable timeline within the budget. Can I start with a free sample? Waiting for your response in chat! Best Regards.
$140 USD in 3 days
5.4
5.4

I appreciate the opportunity to work on the compilation of a comprehensive weather dataset for the entire People’s Republic of China. This project aligns seamlessly with my expertise in data acquisition and management. I will ensure the dataset includes daily minimum and maximum temperatures, precipitation totals, relative humidity, wind speeds, as well as additional parameters from publicly accessible sources, enhancing the dataset's robustness. The data will be meticulously organized in tidy CSV files, with all mandatory variables clearly structured. I will provide a README that outlines the data sources, retrieval methods, whether through API, scraping, or bulk download, along with any time zone considerations and unit conversions. Completing this project also involves providing a detailed data dictionary for clarity. Additionally, a Python or R script will accompany the files, enabling reproducibility of the data pull or future updates. I understand the importance of continuous dates and will explicitly flag any missing days while ensuring that the appropriate numeric types are used to maintain data integrity. I look forward to collaborating on this insightful project and will ensure the highest quality of deliverables. How do you envision the data update process to work in the future?
$30 USD in 11 days
5.2
5.2

A structured, research-grade weather dataset compilation for China (2008-present) delivered as tidy CSV with consistent schemas. Scope: daily Tmin/Tmax for all available mainland provinces plus Hong Kong and Macau, with precipitation, relative humidity, wind speeds, and additional parameters from publicly accessible sources where available. Include continuous dates with explicit gap flags, numeric consistency, and station metadata (lat/lon/elevation/name/province) alongside measurements. Deliverables: 1) CSV files (single national or logically split while keeping identical columns). 2) README detailing data sources, retrieval approach (API/scrape/bulk), time zone handling, unit conversions, and any substitution logic. 3) Data dictionary mapping every column to definition and units. 4) Reproducible Python/R script that refreshes/updates the dataset and validates continuity. Quality & compliance: only publicly licensed/shareable sources suitable for academic research; validation checks for missing days, type integrity, and unit normalization prior to handoff.
$30 USD in 3 days
5.0
5.0

Hi, I got that you are looking for a comprehensive weather dataset covering mainland China, Hong Kong, and Macau from 2008 to present, including daily temperature, precipitation, humidity, wind speed, and other relevant parameters. This is what I can help you with, let's chat. My approach is to utilize advanced data scraping techniques to compile the required data from multiple reliable sources. By leveraging Python scripts for automation, I will ensure a seamless extraction process, maintaining data integrity and quality. You can expect a detailed README file, a comprehensive data dictionary, and a reproducible data pull script for future updates. As final deliverables, you will receive well-structured CSV files containing daily weather data, along with metadata for each station, ensuring a complete and research-ready dataset. One thing I'd like to confirm before we start: Are there any specific regions or time periods that require special attention or may present challenges in data retrieval? Looking forward to discussing further details with you. Regards, Imran
$90 USD in 1 day
5.0
5.0

Hi there, Thank you for outlining such a clear and interesting project. I understand you are seeking a comprehensive, research-ready weather dataset spanning all of China—including Hong Kong and Macau—from 2008 to the present, with daily granularity and a rich set of meteorological variables. You’ve emphasized the importance of data completeness, transparency in sourcing and processing, and academic-grade documentation. I have extensive experience in large-scale data collection and processing, particularly with meteorological and geospatial datasets. My background in Python-based data engineering and data management ensures I can handle the complexity and volume involved. I have previously built similar datasets drawing from sources like NOAA, ECMWF, and China’s CMA, always prioritizing data integrity and reproducibility. For your project, I propose a systematic approach: - Identify and evaluate all reputable, publicly licensed data sources covering the required territory and period. - Develop automated scripts to retrieve, clean, and unify daily weather variables (min/max temperature, precipitation, humidity, wind, and any available extras) for each station or grid point, ensuring consistent units and formats. - Explicitly flag any data gaps and include comprehensive station metadata. - Deliver the dataset as tidy, well-structured CSV files, accompanied by a clear README, a detailed data dictionary, and a Python script to enable future updates or source verification. If any region or period proves challenging, I’ll communicate promptly to suggest alternatives or workarounds. I’m confident my attention to detail and data stewardship can provide you with a robust, research-ready resource. Looking forward to collaborating and answering any further questions you may have!
$140 USD in 5 days
4.6
4.6

Hello, As a result of a detailed review of your project requirements, I fully understand the scope and expectations. I have experience with Python, large-scale data collection, API/bulk-download pipelines, data cleaning, validation, CSV generation, metadata handling, and reproducible research datasets. In my opinion, the key challenge is maintaining a consistent China-wide schema from 2008 to present while combining station/grid data from public sources and explicitly identifying gaps instead of silently filling them. I would build a reproducible Python pipeline that collects daily Tmin/Tmax plus precipitation, humidity, wind, and any additional available variables, normalizes units/time zones, joins station metadata, checks date continuity, and flags missing observations. Depending on volume, I would split outputs by province/year while keeping one consistent column structure. Deliverables will include tidy CSV files, station metadata, README, full data dictionary, validation summary, and an update script so future dates can be appended easily. I have a couple of quick questions: • Do you prefer station-based observations, gridded data, or both where available? • Is there a maximum acceptable CSV/file size for delivery? I’m available to start immediately. Best regards, Carlos.
$30 USD in 7 days
4.3
4.3

Hi, I hope you're doing well. I have carefully reviewed your project, China Weather Dataset Compilation 2008-Now , and I'm confident I can deliver a high quality solution tailored to your requirements. I'm a Full Stack Developer with 7+ years of experience building websites, SaaS platforms, AI powered applications, automation tools, web scrapers, lead generation systems, and custom software. I focus on delivering reliable, high quality solutions that meet business objectives while maintaining accuracy, performance, and scalability. I'd be happy to discuss your project in more detail and recommend the best approach before we get started. I look forward to working with you. Best regards, Abdul Salam Full Stack Developer | Technical Fixes | AI Automation | Lead Generation & Extraction Expert | Websites Dev
$30 USD in 7 days
2.8
2.8

Hello The hardest part is usually aligning data quality, evaluation, and production monitoring with real-world latency and compliance constraints. Integration with existing systems and clear ground truth often drive most of the early risk. What does the current stack look like for ingestion and deployment? Are there SLAs or compliance boundaries I should plan around? Is the work batch-only or does it require real-time inference? Looking forward to learning more about the architecture.
$135 USD in 7 days
0.0
0.0

Hello, I can compile a research-ready China weather dataset from 1 Jan 2008 to delivery day, covering mainland China plus Hong Kong and Macau. I’ll prioritize daily min/max temperature for all available stations or grid points, and add precipitation, humidity, wind, and any other publicly accessible variables I can source. You’ll receive tidy CSVs with a consistent schema, station metadata, a clear README, a data dictionary, and a Python or R script for reproducible updates. I’ll keep dates continuous, flag any missing days explicitly, and document all sources, units, time zones, and conversions. If any regions or periods require substitutions, I’ll flag them early so we can choose the best workaround. Best, Panagiotis
$155 USD in 2 days
0.0
0.0

Hi there, I am a Full Stack Software Engineer with experience in data collection and processing. My technical background allows me to successfully compile a comprehensive weather dataset for China, ensuring accuracy and completeness. This project is crucial for research and analysis of climate trends in China. I will utilize Python for data scraping and ensure all mandatory and additional variables are captured in tidy CSV files. Continuous date records and proper metadata will be prioritized, and I will provide a detailed README and data dictionary for clarity. My systematic approach guarantees a high-quality dataset ready for academic use. Please send a message so we can discuss the details further. Looking forward to working with you. Thank you, Andre
$52 USD in 3 days
0.0
0.0

Hi there, The challenge here is ensuring data completeness across various stations without gaps, while also managing the diverse sources of weather data. Prioritizing the retrieval methods is key to obtaining accurate and consistent records. I propose to gather the mandatory daily temperature data, along with the additional variables you specified, using a combination of API access and web scraping as necessary. I can deliver everything in tidy CSV format, structured logically, and include a README detailing the sources and methods used, as well as a data dictionary for clarity. Would you prefer the data split by province or year for easier management? Looking forward to discussing the details in chat.
$140 USD in 7 days
0.0
0.0

Hi, I'd pull this from NOAA's GHCN-Daily and China's CMA-fed station network rather than scraping individual weather sites, since that gets you consistent daily min/max temps plus the extra variables back to 2008 without gaps in coverage across mainland, Hong Kong, and Macau. I'll write a Python script that fetches everything, checks each date range for missing days and flags them instead of silently skipping, then outputs clean CSVs split by province with a matching station metadata file. You'll also get a plain-language README and a data dictionary so anyone on your team can pick this up later. One thing to watch for: some remote western provinces have thinner station density, so I'll note where coverage gets sparse rather than let it look complete when it isn't. This kind of data pull is something I've done a fair bit of, and untangling unit mismatches between sources is honestly the part I find satisfying. I can have the full package ready in a day. Ready to start whenever you give the word. Best, Emrah
$118 USD in 1 day
0.0
0.0

Transparency, quality and professionalism are the key tenets of my work, which is why I believe I'm perfect for this project. With my extensive experience in data scraping and extraction, your need for a comprehensive weather dataset from China can be accomplished. I understand the importance of accuracy in your project description. I will provide you with a cleaned, thoroughly sourced CSV file; a README that explicitly outlines data sources and retrieval methods; a meticulous data dictionary to ensure column clarity, and a script to update the dataset in future. My web app development expertise significantly complements this project as it requires pulling and managing extensive amounts of information while maintaining top-notch performance, an area on which I thrive. Additionally, my prior experience in SEO set up can be leveraged to improve the digital visibility of your dataset. To sum it up, you won't just be receiving a reliable weather dataset but rather an entire solution from someone who places utmost importance on precision, meeting strict timelines, and ensuring client satisfaction. Let's collaborate and deliver a product that reflects both your vision and the value you're seeking.
$30 USD in 1 day
4.1
4.1

Hello, I'm Rohaan, a full-stack developer, digital marketer, and design specialist with 5+ years of experience and 145+ projects completed. I specialize in Python for data scraping, analysis, and processing. I have a thorough understanding of your requirement for compiling a comprehensive weather dataset for China from 2008 to present. I will meticulously gather daily minimum/maximum temperatures, precipitation totals, relative humidity, wind speeds, and additional parameters from publicly accessible sources. The data will be organized in tidy CSV files, accompanied by a README, data dictionary, and a script for future updates. Let's discuss this project further in chat to ensure a seamless delivery. Best regards, Rohaan
$30 USD in 7 days
0.0
0.0

Hi, I’ve worked with large Python-based data collection and processing pipelines, including cleaning, normalizing, validating, and exporting time-series datasets. For a research dataset covering this many years and locations, I’d prioritize source reliability and reproducibility rather than simply combining files from different providers. Which geographic resolution do you prefer when station observations are unavailable: gridded data as a fallback, or station-only records? I’d first identify publicly shareable sources covering mainland China, Hong Kong, and Macau, then build a reproducible Python pipeline to retrieve and normalize the records into a consistent schema. I’d validate dates, units, station metadata, duplicates, and missing periods, while explicitly flagging gaps instead of silently filling them. The final package would include tidy CSV files, metadata, a data dictionary, README, and an update script so the dataset can be refreshed later. Abel
$140 USD in 7 days
0.0
0.0

Hi, "China Weather Dataset Compilation 2008-Now" looks like exactly what I do. I'm a senior engineer (11+ yrs) specialized in web scraping & data pipelines (Python, anti-bot handling, clean structured output to CSV/JSON/DB) — exactly what this needs. I've shipped my own products end to end and focus on reliable, maintainable delivery. How I'd approach it: 1) Quick chat to lock the exact scope and edge cases 2) Build in small, testable steps so you see progress early 3) Clean handover with docs Ready to begin immediately; I work async and communicate via chat. Ask me anything. Best, Chris
$122 USD in 6 days
0.0
0.0

Hi, I can build the China-wide weather dataset pipeline in Python with consistent CSV output and reproducible updates. Delivering a research-ready record from 1 January 2008 through handover, with temperatures, precipitation, humidity, wind, metadata, and explicit missing-data flags is the goal. I’ve worked on projects where large public datasets needed careful source validation, unit normalization, continuity checks, station metadata, and repeatable collection scripts. I can structure the delivery by province/year if needed, keep one schema throughout, and include the README, data dictionary, and update script. I’m keen to win this project and confident I can deliver a clean, well-documented dataset within the agreed timeline if awarded. Best regards.
$140 USD in 7 days
0.0
0.0

Hong Kong
Payment method verified
Member since Aug 6, 2026
$750-1500 USD
₹750-1250 INR / hour
$15-25 USD / hour
₹12500-37500 INR
₹100-400 INR / hour
₹750-1250 INR / hour
₹750-1250 INR / hour
₹750-1250 INR / hour
₹750-1250 INR / hour
₹12500-37500 INR
₹75000-150000 INR
₹37500-75000 INR
$250-750 USD
₹400-750 INR / hour
$10 USD
₹12500-37500 INR
₹1500-12500 INR
₹1500-12500 INR
$15-25 USD / hour
$15-25 USD / hour
₹12500-37500 INR