Freelance Data Consultant

Freelance Data Consultant completes the practical tasks listed here and this page provides 23 real tasks for reference. The examples show who the role serves and how outcomes are tracked. Each one shows where we found it, and comes with an AI prompt you can copy and use straight away.

23evidenced tasks
23ready prompts
7tools of the trade
15-2051.00O*NET-SOC code
262,440hold this job (US, BLS 2025)
$120,230median pay/yr (US)
Open Freelance Data Consultant in the interactive atlas →

What it pays

Government survey numbers — not estimates, not ads.

Half of all Data Scientists in the U.S. earn more than $120,230 a year — the middle 80% land between $67,240 and $199,130. About 262,440 people in the U.S. do this work. Figures are for the U.S. occupation group “Data Scientists”. (U.S. Bureau of Labor Statistics survey, published 2025.) In India, Professionals earn about ₹38,298 a month on average — around ₹4.6 lakh a year (government PLFS survey via ILOSTAT, occupation-family figure).
$120,230typical pay / year
262,440people in this work
$199,130+top 10% earn
₹4.6 lakha year in India (family avg)
Think you get this job?Six quick questions on how it really works — with a hint and the reason behind every answer.
Test yourself →

The work, task by task

These are the real jobs-to-be-done, not a wish list. Each task shows where we found it, and the prompt underneath is written for that exact task.

Learning3

Manage large amounts of data

+
Ingest the six quarterly ICT export files, deduplicate and standardise fields, flag missing supplier IDs,…
Ingest the six quarterly ICT export files, deduplicate and standardise fields, flag missing supplier IDs, then publish the cleaned master dataset ready for downstream modelling by Tuesday lunch so I can hand it to clients without manual fixes.
The tools that do the workApache SparkESCOO*NETWikipedia

Develop and test data models

+
Build three candidate schema variants for the recommender training set from the cleaned master table, run…
Build three candidate schema variants for the recommender training set from the cleaned master table, run k-fold validation on each, capture performance metrics and a short recommendation note so we can pick the production model by Friday.
The tools that do the workApache SparkESCOjob descriptionsWikipedia

Apply machine learning techniques

+
Experiment with three supervised learning approaches on the cleaned training set, track hyperparameters and…
Experiment with three supervised learning approaches on the cleaned training set, track hyperparameters and feature importances, compare ROC and precision-recall curves, and save the best pipeline with version notes for deployment review next Tuesday.
The tools that do the workApache Sparkjob descriptionsO*NET
Keeping the record2

Present findings through reports and presentations

+
Draft a ten-slide report summarising key patterns in ICT usage, include top three actionable recommendations,…
Draft a ten-slide report summarising key patterns in ICT usage, include top three actionable recommendations, attach dataset snapshots and reproducible code snippets, and export a presentation for the client meeting on Monday morning.
The tools that do the workAtlassian ConfluenceESCOjob descriptionsO*NET

Communicate insights to stakeholders

+
Write a one-page stakeholder brief that explains model uplift, expected operational impacts, and three…
Write a one-page stakeholder brief that explains model uplift, expected operational impacts, and three implementation risks with mitigation owners, then circulate it to Priya in procurement and Ahmed in ops before Wednesday standup.
The tools that do the workAtlassian ConfluenceESCOjob descriptionsO*NET
Move1

Merge data sources

+
Join the CRM export, transaction ledger, and supplier registry into a single canonical table, reconcile…
Join the CRM export, transaction ledger, and supplier registry into a single canonical table, reconcile conflicting IDs, create provenance flags, and produce a mapping file so analysts can trust joins for the monthly dashboard.
The tools that do the workApache SparkESCOjob descriptions
The daily work17

Identify business problems and data solutions

+
Map our three revenue leaks and the decisions that keep them open, list what evidence I need from sales,…
Map our three revenue leaks and the decisions that keep them open, list what evidence I need from sales, support and product to prove each leak exists, and propose two data solutions I can build in two weeks to stop the biggest one.
The tools that do the workApache AirflowAtlassian JIRAjob descriptionsO*NET

Analyze data to identify patterns and trends

+
Find patterns in the last 18 months of customer usage, churn and support tickets, summarise the three…
Find patterns in the last 18 months of customer usage, churn and support tickets, summarise the three strongest signals that predict churn, and deliver the SQL and validation tests I used so the product team can reproduce them.
The tools that do the workApache Sparkjob descriptionsO*NET

Categorize and organize data

+
Take the raw client, transaction and metadata files, standardise names and IDs, produce a clean schema with…
Take the raw client, transaction and metadata files, standardise names and IDs, produce a clean schema with three categorical taxonomies and a data dictionary, and hand over the organized tables ready for analysis.
The tools that do the workAlteryxjob descriptionsWikipedia
km/h RPM

Create data visualizations and dashboards

+
Build three executive dashboards: customer health by cohort, weekly acquisition funnel, and feature adoption…
Build three executive dashboards: customer health by cohort, weekly acquisition funnel, and feature adoption heatmap; include clear metric definitions, refresh cadence, and one exportable CSV per chart.
The tools that do the workApache HiveESCOjob descriptions

Apply sampling techniques to determine groups to be surveyed or use complete enumeration methods.

+
Decide whether to sample or enumerate our 120,000 users for the satisfaction survey, show the confidence and…
Decide whether to sample or enumerate our 120,000 users for the satisfaction survey, show the confidence and cost tradeoffs for three sampling strategies, and pick the smallest sample that meets 95 percent confidence.
The tools that do the workAmazon Elastic Compute Cloud EC2O*NET

Design surveys, opinion polls, or other instruments to collect data.

+
Draft the customer survey with clear demographics, two validated satisfaction scales, three behaviour…
Draft the customer survey with clear demographics, two validated satisfaction scales, three behaviour questions, and the routing logic so we get complete responses from support cases in the next sprint.
The tools that do the workAtlassian ConfluenceO*NET

Combine statistical knowledge with coding

+
Combine the cleaned ICT logs, user ratings and product metadata into a reproducible pipeline that trains a…
Combine the cleaned ICT logs, user ratings and product metadata into a reproducible pipeline that trains a collaborative recommender, validates performance with holdout folds, and outputs model artifacts and a short README for deployment by Friday.
The tools that do the workApache SparkWikipedia

Visualize data to identify patterns

+
Produce a set of interactive charts and a one-page dashboard that surfaces usage patterns, seasonality and…
Produce a set of interactive charts and a one-page dashboard that surfaces usage patterns, seasonality and outliers from the ICT survey and event data, annotated with brief executive takeaways for Monday's client review.
The tools that do the workAlteryxWikipedia

Find and interpret rich data sources

+
Locate diverse high-quality sources — public ICT indicators, partner CSV exports, and internal logs —…
Locate diverse high-quality sources — public ICT indicators, partner CSV exports, and internal logs — document access steps and licensing, then deliver a catalog of normalized datasets with richness and reliability notes by Wednesday.
The tools that do the workApache HiveESCO

Ensure consistency of data-sets

+
Standardize the three client datasets, reconcile schema and coding differences, implement row-level…
Standardize the three client datasets, reconcile schema and coding differences, implement row-level deduplication and validation rules, and publish the cleaned canonical table with checksums and a change log before handoff.
The tools that do the workAlteryxESCO

Recommend ways to apply the data

+
Review the cleaned ICT and engagement datasets, write three practical recommendations for product…
Review the cleaned ICT and engagement datasets, write three practical recommendations for product personalization, retention tactics and a lightweight A/B test plan, and attach expected metrics and effort estimates for next sprint.
The tools that do the workApache SparkESCO

Read scientific articles, conference papers, or other sources of research to identify emerging analytic trends and technologies.

+
Read recent journal articles and conference papers on recommender systems and data infrastructure from the…
Read recent journal articles and conference papers on recommender systems and data infrastructure from the last 18 months, extract emerging analytic methods and toolchains, and produce a two-page briefing with implications for clients building recommender engines.
The tools that do the workApache SparkO*NET

Determine available and useful data for projects

+
Inventory internal and public data sources for the client’s pilot, list available fields, update frequency,…
Inventory internal and public data sources for the client’s pilot, list available fields, update frequency, access method, and likely quality issues, then recommend which sources to use first to train a basic recommender prototype.
The tools that do the workApache Hivejob descriptions

Clean and process raw data

+
Take the raw CSV exports from the pilot, profile missingness and duplicates, apply cleansing rules to…
Take the raw CSV exports from the pilot, profile missingness and duplicates, apply cleansing rules to standardise identifiers and timestamps, and deliver a versioned, documented dataset ready for feature engineering.
The tools that do the workAlteryxjob descriptions

Interpret data analysis results

+
Translate the model outputs and evaluation metrics into a client-facing one-page summary that explains lift,…
Translate the model outputs and evaluation metrics into a client-facing one-page summary that explains lift, failure modes, and recommended next experiments, with a slide that shows three actionable decisions for product and engineering.
The tools that do the workAtlassian Confluencejob descriptions

Monitor and improve model performance

+
Set up the monitoring plan for the recommender: define performance metrics, alert thresholds, data drift…
Set up the monitoring plan for the recommender: define performance metrics, alert thresholds, data drift checks, and a weekly report schedule, then run the first week of baselines and log any anomalies and corrective actions.
The tools that do the workApache Airflowjob descriptions

Collaborate with cross-functional teams

+
Organise a cross-functional kickoff with product, engineering, and UX: circulate an agenda listing data…
Organise a cross-functional kickoff with product, engineering, and UX: circulate an agenda listing data needs, integration constraints, and success criteria, capture action owners and deadlines, and circulate meeting notes with next steps.
The tools that do the workAtlassian JIRAAtlassian Confluencejob descriptions

Says who?

These are the pages we read to build this. Open any of them and check us.

The logs, files & records this job keeps

Shared with other careers — the same record means something different in each.

Related careers

Same family of work — each with its own tasks and prompts.

The LLOS Work Atlas is the world's largest evidenced task library — a map of human work, with a ready prompt behind every task. 1,774 careers · every task named by the sources that witnessed it — O*NET, ESCO, real job descriptions, Wikipedia — and the deepest tasks by several at once. And it is honest about limits: where AI cannot help, the map says so.

The rest of the map

Same library, five ways in.

Copyright © LLOS.ai · 2026 — original pedagogy, voice, and design — all rights reserved.
Built on public evidence: O*NET®, ESCO, Wikipedia, U.S. Bureau of Labor Statistics, ILOSTAT. All sources & licenses