Senior RnD Data Analyst
Posted 21 days ago · 0 applicants
Saving, applying or scoring takes a few seconds to set up your free account.
The role in plain words
This role involves taking full ownership of data models, data flows, and validation processes across the R&D department. You will perform in-depth data quality assessments, develop monitoring frameworks, and curate golden datasets used for machine learning model training. Additionally, you will collaborate closely with data scientists, engineers, and product teams to ensure data correctness and reliability.
- 5+ years of hands-on experience in data analysis, data operations, analytics engineering, or a similar data-focused role
- Strong analytical skills
- High proficiency in SQL
- High proficiency in Python for data investigation, validation, and dataset creation
- Experience building and maintaining golden datasets
- Familiarity with tools such as Spark, Airflow, or similar technologies
Extracted from the job description · kept up to date automatically
Who this suits
This role suits experienced data professionals with over five years of hands-on experience who are highly proficient in SQL and Python. It is less ideal for those without experience in building golden datasets, managing annotation processes, or working with modern data pipelines and cloud environments.
Full job description
Original listing · kept for referenceWe’re looking for a brilliant, hands-on, and mission-driven R&D Data Analyst to join our team at IVIX. Our technology helps governments uncover hidden business activity and fight financial crimes by transforming massive volumes of public web data into actionable intelligence.
In this role, you will own the data end-to-end — from deeply understanding the business flow, to validating pipelines, to ensuring the accuracy, reliability, and usability of the data powering our product. You’ll work closely with engineering, data science, and product teams to ensure our data is technically sound, business-aligned, and ready for analytical and machine learning use cases.
IVIX is the first AI-powered solution designed specifically to address a $20 trillion problem: illuminating the global shadow economy. IVIX leverages Open-Source Intelligence (OSINT) and highly advanced, cutting-edge technologies to reveal illicit business activity around the world, empowering governments in their mission to fight financial crime and close the tax gap.
IVIX employs a variety of AI tools (deep neural networks, large language models, and predictive modeling) as well as advanced data analytics to rapidly pinpoint large-scale illicit business activity, so government authorities can combat financial crime in the digital age.
Led by security, tech, tax and financial crime experts, and advised by a diverse team of former IRS commissioners, IVIX works with dozens of state and federal governments globally.
• Take full ownership of the data across R&D — including data models, data flows, and the processes that power our analytics and product features.
• Perform in-depth data validation and quality assessments to ensure correctness, consistency, and reliability across pipelines.
• Develop and maintain frameworks and methodologies for data assurance, anomaly detection, and monitoring.
• Create, manage, and curate golden datasets, ground truth, and annotated datasets used for data science, ML model training, evaluation, and benchmarking.
• Work closely with data scientists to ensure datasets align with modeling needs, experimental design, and feature development.
• Analyze complex datasets to surface insights, edge cases, and discrepancies that impact product decisions or customer outcomes.
• Document data flows, definitions, validation logic, and labeling guidelines while championing data transparency and accessibility across the organization.
• Serve as the subject-matter expert for data correctness and completeness.
• 5+ years of hands-on experience in data analysis, data operations, analytics engineering, or a similar data-focused role.
• Strong analytical skills and high proficiency in SQL and Python for data investigation, validation, and dataset creation.
• Experience building and maintaining golden datasets, creating ground truth, and managing annotation/labeling processes for ML or analytics use cases.
• Familiarity with modern data pipelines, ETL/ELT processes, and data workflow tools.
• Ability to understand complex business processes and translate them into data logic and validation frameworks.
• Experience working with large datasets in cloud environments (AWS preferred).
• Familiarity with tools such as Spark, Airflow, or similar technologies - a strong advantage.
• A passion for data, problem-solving, and getting into the details.
Questions about this role
- This listing did not state a salary. We only show pay when the employer publishes it.
- 5+ years of hands-on experience in data analysis, data operations, analytics engineering, or a similar data-focused role, Strong analytical skills, High proficiency in SQL, High proficiency in Python for data investigation, validation, and dataset creation, Experience building and maintaining golden datasets