Offres d'emploi
Trouvez des postes près de chez vous, sur site, hybrides ou à distance.- Emplois similaires à : AI Curation Data Scientist
AI Curation Data Scientist
xCuresNew YorkThe xCures platform helps improve clinical care via comprehensive, intelligent access to healthcare data on an AI-assisted platform. Delivered using a Software as a Service (SaaS) model, xCures enable
AI Curation Data Scientist
Recruit GroupNew YorkWhat You’ll DoDevelop and optimize software pipelines for extracting and integrating structured and unstructured healthcare dataBuild and maintain AI/ML workflows for data classification, normalizatio
Data Scientist
SOLACE HEALTH LLCNew YorkAbout Solace Healthcare in the U.S. is fundamentally broken. The system is so complex that 88% of U.S. adults do not have the health literacy necessary to navigate it without help. Solace cuts through
Data Scientist
SleeperNew YorkSleeper is seeking a sharp, curious, and collaborative Data Scientist to join our growing team. This role will partner closely with product, engineering, and business teams to drive decision-making th
Data Scientist
SystematixNew YorkABOUT THE PROJECT Our client is seeking an experienced Data Scientist to contribute to a cutting-edge data science initiative focused on advanced optimization and predictive modeling. The role support
Data Scientist
AdelaideNew YorkTL;DR Fast-growing ad-measurement company looking for an individual to work with our data science team to help build the AU and drive our measurement product forward.Who we are: Adelaide is the leader
Data Scientist
Spirent CommunicationsNew YorkJoin to apply for the Data Scientist role at Spirent Communications Join to apply for the Data Scientist role at Spirent Communications Get AI-powered advice on this job and more exclusive features. S
Data Scientist
KubeltNew YorkResponsibilities Collect, process, and analyze data to deliver insights and business growth Develop predictive models, algorithms, and visualizations based on complex network activities Analyze and va
Data Scientist
Avenue CodeNew YorkAvenue Code is the leading software consultancy focused on delivering end-to-end development solutions for digital transformation across every vertical. We’re privately held, profitable, and have been
Data Scientist
Mackin TalentNew YorkData Scientist to support the Authentication Services team within Mobile Identity. The role focuses on modernising our authentication stack to enable client's sustainable growth through high quality U
Senior Data Scientist
DuckDuckGoNew YorkWho We Are Hi, we're DuckDuckGo, the online protection company and remote-first team of 300+ on a mission to raise the standard of trust online. Founded in 2008 and profitable since 2014, annual reven
Staff Data Scientist
OpenXNew YorkCompany at a Glance OpenX is focused on unleashing the full economic potential of digital media companies. We do this by making digital advertising markets and technologies that are designed to delive
Data Analyst Associate Data Scientist
Framework VenturesNew YorkOverview Integra is hiring entrepreneurial and quantitatively-skilled individuals with data analytics experience who are interested in applying their skills to identify, investigate and analyze fraud
Operations Data Scientist
Aspect SoftwareNew YorkThis range is provided by Aspect Software. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $115,000.00/yr - $120,000.00/yr Direct m
Senior Data Scientist
SOSiNew YorkFounded in 1989, SOSi is among the largest private, founder-owned technology and services integrators in the defense and government services industry. We deliver tailored solutions, tested leadership,
Senior Data Scientist
Revelio LabsNew YorkWho We Are Revelio Labs provides workforce intelligence. We absorb and standardize hundreds of millions of public employment records to create the world’s first universal HR database, allowing us to s
Data Scientist- Platform Integrity
SleeperNew YorkAbout Sleeper Sleeper is a sports-focused games platform with messaging at its core. We are a young and energetic company, fueled by a passion for sports and a drive for innovation. Our mission is to
Senior Data Scientist, Consumer
RedditNew YorkLocation: US remote-friendly or any office location - SF, LA, CHI, NYThe Data Science Team at Reddit is growing and we are looking for experienced Data Scientists to partner with our cross‑functional
Senior Data Scientist | Remote
Crossing HurdlesNew YorkPosition: Experienced & Credentialed Data ScientistsType: Hourly contractCompensation: $100-$160 per hourLocation: RemoteCommitment: 10–40 hours/weekRole ResponsibilitiesWork on projects that focus on
Staff Data Scientist - RiskOS
Socure IncNew YorkWhy Socure? Socure is building the identity trust infrastructure for the digital economy — verifying 100% of good identities in real time and stopping fraud before it starts. The mission is big, the p
Senior Data Scientist III
RELXNew YorkAbout the Team LexisNexis Legal & Professional, which serves customers in more than 150 countries with 11,800 employees worldwide, is part of RELX, a global provider of information‑based analytics and
Junior Data Scientist Healthcare
Framework VenturesNew YorkIntegra Med Analytics, based in Austin, TX, is a team of researchers and economists who seek to use forensic data analysis to promote integrity in the healthcare system. We are looking for strong cand
Staff Forecasting Data Scientist
Omada HealthNew YorkJob Overview Omada Health is looking for a Staff Forecast Data Scientist to lead the technical development and automation of our enrollment forecasting capability. This role will build and scale forec
Senior Marketing Data Scientist
MozillaNew YorkWhy Mozilla? Mozillians design, build and distribute open‑source software that enables people to enjoy the internet on their terms.About this team and role The Marketing Data Science team sits within
Senior Data Scientist, Analytics
TRMNew YorkBuild a Safer World. TRM Labs provides blockchain analytics and AI solutions to help law enforcement and national security agencies, financial institutions, and cryptocurrency businesses detect, inves
AI Curation Data Scientist
- New York, New York, United States
- New York, New York, United States
À propos
About the role: Reporting to the VP of Data Science, the AI Curation Data Scientist will, using traditional computing and custom AI model training, work on mission-critical projects driven by xCure’s product development needs and will expand xCures’ complex, innovative health data processing, extraction, and analysis capabilities. We’re looking for an individual who values team-building, cooperation, and communications with colleagues to serve the needs of our customers. Projects will include significant data processing challenges, such as C-CDA XML parsing and de-identification of structured and unstructured EHR content. Equally important projects will address data curation and custom AI model training. You will author software and AI models and contribute to data set curation. You will coordinate data set quality assurance within an innovative, fast-moving team. Team responsibilities are key requirements for this position, which will deliver large and complex data products and data analysis tools.
This position is fully remote, but will coordinate very closely with a small team and thus requires excellent communication and coordination skills. Occasional travel is required.
This job is right for you if you like:
A high-energy start-up working with a brilliant and passionate team
Working on problems that make a real difference in people’s lives
Understanding and delivering on reliable and well-characterized products and deliverables within a highly innovative and fast-changing environment: clinical data extraction and aggregation, the relation between data processing and QA framework; LLM tuning and training, pedantic data curation, compute architecture, data exchange.
Rockstar teammates: you will be working with a strong team with decades of prior work experience in artificial intelligence, software systems, molecular biology, and clinical medicine
Innovation and problem solving to provide order-of-magnitude improvements in capabilities for data handling and analysis while maintaining traceable data and methods development
Responsibilities:
Developing and testing data extraction and integration software for structured EHR content (XML, FHIR) and unstructured text content (attached documents)
Organizing and contributing to data set curation for model training
Tuning and training LLMs
Maintaining a strong understanding of PHI/PII and de-identification policies and strategies at xCures and implementing software solutions compliant with policies and strategies
Developing and implementing tests of data extraction and aggregation performance to improve efficiency, timeliness, and cost-effectiveness
Implementing and maintaining code repositories
Working closely with manager to explore methods, test hypotheses, and collaboratively implement innovative solutions for data science
Coordinating as required for a fully remote role
Working with Engineering and other groups to improve overall company efficiency and effectiveness
Required Skills and Qualifications:
Masters degree or equivalent experience in Computer Science, Software Engineering, Statistics, Biology, or related field
Minimum of 5 years of hands‑on experience in data science, machine learning, AI, data analysis, software development, and/or predictive analytics
Experience applying generative AI and transformer models, especially training of LLMs
Significant experience with curating data sets to train LLMs
Significant hands‑on coding experience with LLMs, embeddings models, sentence_transformers, and authoring python code to build data extraction and/or classification tools
Significant prior work experience with parsing XML, JSON, and/or other complex data formats, preferably C-CDA health data
Experience with TensorFlow, PyTorch, and/or scikit‑learn
Software development skills including git
Proven efficiency using, and cautious approach to using, LLM-assisted coding
Experience writing unit and integration tests for scientific/clinical data software as well as with developing scientifically motivated data quality assessments
Flexible, innovative, can‑do approach to delivering software and data products balanced with team cooperation
A passion for successful delivery of team work products
Must reside in the United States.
Must have authorization to work in the United States.
Preferred Skills and Qualifications:
Extensive experience with data handling efficiency tools, such as jq, xq, Unix command-line tools such as sed, bash programming
Deep understanding of regex
Extensive AWS experience and understanding of tradeoffs for different types of data storage for AI training
Significant experience with PHI and PII, HIPAA, and de-identification is a major plus
Software development experience in multiple coding languages
Confidence extending the capabilities of open source tools
Experience with multiple approaches to LLM-assisted coding, such as within Visual Studio, copilot, Claude Code; and familiarity with frontier and open model capabilities
Experience with remote teams and solving technical project communications challenges
Notes: This is a big list. Don’t worry if you do not meet every qualification or wishlist item. If you are passionate, ambitious, adept, and mission-aligned, then we want to hear from you — even if you don’t check every box listed here. True talent shines through and transcends a list of bullet points. To apply, please send your cover letter and resume to ds-jobs@xcures.com
Salary range : 100K to 165K annually
401k
xCures acknowledges that equal opportunity for all persons is a fundamental human value. Each employee and applicant will be considered on the basis of individual ability and merit, without regard to race, color, religion, age, sex, sexual orientation, gender identity, gender expression, pregnancy, national origin, marital status, physical disability, mental disability, medical condition, genetic information, protected military or veteran status, or any other characteristics.
#J-18808-Ljbffr
Compétences linguistiques
- English
Cette offre provient d’une plateforme partenaire de TieTalent. Cliquez sur « Postuler maintenant » pour soumettre votre candidature directement sur leur site.