Jobedly Post a Job

Data Scientist — Blockchain Intelligence

Merkle Science · Remote
RemoteFull-timeFinance & InsuranceMid Level$111,000–$150,000/yr
Apply on Jobedly ⚡ One-click AI Apply

About the Data Scientist Blockchain Intelligence role

Data Scientist Blockchain Intelligence positions focus on delivering results in their domain. This page aggregates open Data Scientist Blockchain Intelligence roles and what employers typically expect.

**⚡️ About Merkle Science** Merkle Science provides blockchain transaction monitoring and intelligence solutions for web3 companies, digital asset service providers, financial institutions, law enforcement and government agencies to detect, investigate, and prevent illicit use of cryptocurrencies. Our vision is to make cryptocurrencies safe and provide infrastructure for the safe and compliant growth of cryptocurrencies. Merkle Science is headquartered in New York with offices in Singapore, Bangalore and London. The team has combined experience across Bank of America, Paypal, Luno, Thomson Reuters and Amazon. The company has raised over $27M from SIG, Beco, Republic, DCG, Kenetic, GGV and several others. ## About the role We turn raw on-chain activity into trustworthy intelligence — clustering addresses into real-world entities, attributing them to services and actors, and surfacing risk for compliance and investigations teams. We're looking for a data scientist who is as comfortable shipping a heuristic to production as they are designing it: someone who can move from a messy hypothesis to a working pipeline without waiting on someone else to wire up the data. You'll work closely with our attribution and clustering leads on models and heuristics that run across billions of transactions and multiple chains (Bitcoin, Ethereum, Tron, Solana, and more). ## What you'll do - Design, test, and ship clustering and attribution heuristics, and measure them with real precision/coverage metrics rather than vibes. - Own your data end to end — pull, clean, join, and model large on-chain datasets without depending on a separate team for every query. - Build and maintain the pipelines that take a heuristic from notebook to production, including backfills, incremental runs, and validation. - Investigate edge cases (mixers, bridges, exchange hot wallets, consolidation patterns) and translate findings into repeatable logic. - Partner with investigations and product to define what "correct" looks like and benchmark against ground truth. - Prototype quickly, then harden what works. ## What we're looking for - 4+ years building data science or data engineering systems that actually shipped (not just notebooks). - Strong Python and SQL; comfortable with large datasets and the gotchas of joins, dedup, and skew at scale. - Solid grasp of clustering, graph/network analysis, or entity resolution — and a habit of validating results, not just producing them. - Ability to reason about precision vs. coverage trade-offs and defend your metrics. - Self-directed: you can scope an ambiguous problem, get the data yourself, and drive it to a result. ## Our tech stack You don't need to have used all of these, but here's what you'd be working with day to day: - **Databricks** — our lakehouse and processing backbone. Large-scale on-chain datasets are transformed and modeled here via Spark and SQL; most heuristics run as Databricks jobs against billions of transactions. - **Kafka** — real-time ingestion of on-chain and transaction data. New blocks and events stream in continuously, so a lot of our work is designed to run incrementally rather than as one-off batch jobs. - **Python** — the primary language for everything from exploratory analysis to production heuristics and pipeline code. - **TigerGraph** — our graph database, where addresses, transactions, and entities live as a network. Clustering, traversals, and relationship queries (who funds whom, consolidation paths, entity linkage) happen here. Supporting cast you'll likely touch: - **SQL** everywhere — for ad-hoc analysis, validation, and defining ground-truth datasets. - **Columnar / analytical stores** (e.g., ClickHouse) for fast aggregate queries over large tables. - **Orchestration & scheduling** for backfills and recurring pipeline runs. - **Git / GitHub** for version control and code review — we expect pipelines and heuristics to be reviewed like any other code. - **GCP** as our cloud environment. ##…

Salary estimate

$111,000 – $150,000/yr
Provided by the employer.

Skills for this role

PythonSQLGCPGITSparkKafkaData Science

Resume tips for Data Scientist Blockchain Intelligence applicants

Interview preparation

Prepare concrete STAR-format stories that show Data Scientist Blockchain Intelligence outcomes you drove.

Research the employer's product and recent news before the interview.

Be ready to explain how you'd approach a typical Data Scientist Blockchain Intelligence problem end to end.

Have thoughtful questions ready about the team, tools and success metrics.

About Merkle Science

Merkle Science is actively hiring on Jobedly. Explore their open roles and what it's like to work there.

Apply on Jobedly ⚡ One-click AI Apply

Similar jobs

Companies hiring for similar roles