paidwork.com
Aug 2, 2026
Database dump from PaidWork, a get-paid-to/microtask platform. Contains user accounts, email addresses, personal details (names, birthdays, gender, location data including lat/long), billing information (full names, addresses, bank account numbers, BIC codes, PayPal emails, crypto wallets), withdrawal history with transaction amounts, and mailing list data. Users appear to span globally with significant representation from Poland, Morocco, Algeria, Egypt, and other countries.
Data found in this dataset
Source files
Expand any file to inspect its column headers and the LLM's field-mapping reasoning, recorded during ingestion.
dump__mailing.csv2 columns23,645,231 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
| Source column | Mapped field | Confidence | LLM assessment |
|---|---|---|---|
| 2 | high | header "email" resolves to PII field "email" | |
| 5 | firstName | high | header "firstname" resolves to PII field "firstName" |
Notes: Heuristic auto-detection: header-named PII columns confirmed by data conformance
dump__user_about.csv7 columns0 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
| Source column | Mapped field | Confidence | LLM assessment |
|---|---|---|---|
| 1 | firstName | high | [1] header 'about_firstname', values are common given names like 'Jakub', 'aseel', 'Sherif' |
| 2 | lastName | high | [2] header 'about_lastname', values are surnames like 'Jabłoński', 'mahyoub', 'Mostafa' |
| 3 | gender | high | [3] header 'about_gender', values are gender codes (1/2) |
| 4 | dob | high | [4] header 'about_birthday', values match DD/MM/YYYY date pattern like '20/12/2002' |
| 12 | city | high | [12] header 'about_city', values are city names like 'Warsaw', 'Karachi', 'Agadir' |
| 13 | zip | high | [13] header 'about_postal_code', values are postal codes like '00-000', '12311' |
| 14 | country | high | [14] header 'about_country_code', values are ISO country codes like 'BR', 'EG', 'PK' |
Notes: 29 columns total. Columns 5-8 (about_age, about_age_year, about_age_month, about_age_day) are derived/computed age fields — skipped as non-primary DOB data since column 4 has the canonical birthday. Columns 10-11 (region name/code) are sub-national region data with no direct PII field mapping — skipped. Columns 15-16 (lat/long) are geographic coordinates — skipped. Columns 9, 17-27 are profile images, coded interest/work/education flags, relationship flags, and timestamps — all skipped.
dump__user_billings.csv7 columns3,641,684 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
| Source column | Mapped field | Confidence | LLM assessment |
|---|---|---|---|
| 4 | firstName | high | [4] header 'billings_firstname', values are common given names |
| 5 | lastName | high | [5] header 'billings_lastname', values are common surnames |
| 6 | fullName | high | [6] header 'billings_fullname', values combine first and last names |
| 7 | address1 | high | [7] header 'billings_address1', values are street addresses |
| 8 | address2 | high | [8] header 'billings_address2', values are additional address lines |
| 9 | city | high | [9] header 'billings_city', values are city names |
| 10 | zip | high | [10] header 'billings_postal_code', values are postal codes |
Notes: 15 columns total, 6 contain PII. Columns 0-3, 11-14 are internal IDs, IPs, currency codes, and timestamps — excluded per rules.
dump__user_billings_method_withdrawal.csv7 columns404,257 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
| Source column | Mapped field | Confidence | LLM assessment |
|---|---|---|---|
| 7 | fullName | high | header "fullname" resolves to PII field "fullName" |
| 8 | address1 | high | header "address1" resolves to PII field "address1" |
| 9 | address2 | high | header "address2" resolves to PII field "address2" |
| 10 | city | high | header "city" resolves to PII field "city" |
| 11 | zip | high | header "postal_code" resolves to PII field "zip" |
| 13 | country | high | header "country_code" resolves to PII field "country" |
| 14 | high | header "email" resolves to PII field "email" |
Notes: Heuristic auto-detection: header-named PII columns confirmed by data conformance
dump__user_billings_withdraw.csv7 columns101,084 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
| Source column | Mapped field | Confidence | LLM assessment |
|---|---|---|---|
| 10 | fullName | high | header "fullname" resolves to PII field "fullName" |
| 11 | address1 | high | header "address1" resolves to PII field "address1" |
| 12 | address2 | high | header "address2" resolves to PII field "address2" |
| 13 | city | high | header "city" resolves to PII field "city" |
| 14 | zip | high | header "postal_code" resolves to PII field "zip" |
| 16 | country | high | header "country_code" resolves to PII field "country" |
| 17 | high | header "email" resolves to PII field "email" |
Notes: Heuristic auto-detection: header-named PII columns confirmed by data conformance
dump__user_contact.csv1 column1,804,920 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
| Source column | Mapped field | Confidence | LLM assessment |
|---|---|---|---|
| 1 | phone | high | header "phone_number" resolves to PII field "phone" |
Notes: Heuristic auto-detection: header-named PII columns confirmed by data conformance
dump__user_personal_id.csv0 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
Notes: All three columns are internal identifiers: 'id' and 'user_id' are numeric auto-increment/foreign keys, and 'personal_id' is a structured internal reference code (format NNNN-NNN-NN). No PII fields are present in this file.
dump__users.csv1 column23,257,032 rows
File structure
Format: CSV·Delimiter: Comma·Has header: yes·Quote: "
| Source column | Mapped field | Confidence | LLM assessment |
|---|---|---|---|
| 1 | high | header "email" resolves to PII field "email" |
Notes: Heuristic auto-detection: header-named PII columns confirmed by data conformance