InstantServe LLC

Senior Data Scientist

⭐ - Featured Role | Apply direct with Data Freelance Hub
This role is for a Senior Data Scientist in Woodlawn, MD, with a contract length of unspecified duration. Pay rate is also unspecified. Key skills include NLP, Python, SQL, and experience with Generative AI. A Master's degree and 10+ years of experience are required.
🌎 - Country
United States
πŸ’± - Currency
$ USD
-
πŸ’° - Day rate
Unknown
-
πŸ—“οΈ - Date
August 19, 2026
πŸ•’ - Duration
Unknown
-
🏝️ - Location
On-site
-
πŸ“„ - Contract
Unknown
-
πŸ”’ - Security
Unknown
-
πŸ“ - Location detailed
Woodlawn, MD
-
🧠 - Skills detailed
#Matplotlib #Hadoop #Deployment #Pandas #Airflow #NLP (Natural Language Processing) #Data Engineering #Programming #EC2 #SQL Server #Langchain #Version Control #Impala #Libraries #Database Management #Python #MySQL #NumPy #SpaCy #AWS (Amazon Web Services) #PyTorch #Data Analysis #Mathematics #Oracle #Statistics #DevOps #NLTK (Natural Language Toolkit) #Model Deployment #GIT #SQL (Structured Query Language) #Spark (Apache Spark) #ML (Machine Learning) #Cloud #AI (Artificial Intelligence) #HTML (Hypertext Markup Language) #Scala #TensorFlow #Apache Spark #Leadership #Data Science #Anomaly Detection #Computer Science #Regular Expressions #Monitoring #Web Services #Azure
Role description
β€’ β€’ Selected candidate must be able to obtain and maintain a public trust clearance β€’ β€’ β€’ β€’ Selected candidate must be willing to work on-site in Woodlawn, MD 5 days a week β€’ β€’ ACCEPTING LOCALS only β€’ β€’ Master's and 10+ years of experience, Bachelor's and 12+ years of experience or 18+ years in lieu of a degree β€’ β€’ Key Required Skills β€’ Solid Experience with Natural Language Processing (NLP), Python, NLP frameworks, SQL, Pandas, NLTK and SPACy. β€’ Experience with Generative AI and Large Language Models (LLM) β€’ Excellent Communication skills Position Description β€’ Hands on experience in Python, NLP frameworks, SQL, Pandas, NLTK, SPACy and LLMs β€’ Well versed in SQL and analyzing trends and transactional data. β€’ Understand real world challenges and develop automated data solutions β€’ Develop, test, and deploy new techniques for NLP understanding β€’ Scalable development/deployment of ML and Generative AI approaches (such as Large Language Models (LLMs) β€’ Train and optimize NLP/LLM models and create Python based pipelines β€’ Experience building cloud native solutions on AWS β€’ Determine the nature of analytic problems, evaluate options, and offer recommendations for resolution. β€’ Advise on the methods and data needed and/or available to evaluate the (intelligence or data) problem. β€’ Collaborate with data collectors and analysts to identify and close gaps on complex monitoring problems. β€’ Provide accurate, timely, complex, and sophisticated data analysis. Skills Requirements β€’ Bachelor’s degree in Statistics, Applied Mathematics, Computer Science, or Information Science with industry experience on Python, NLP frameworks, SQL, Pandas, NLTK and SPACy, data science, and AI/ML/LLM engineering. β€’ 15 year s+candidate β€’ Solid Experience with Natural Language Processing (NLP), Python, NLP frameworks, SQL, Pandas, NLTK and SPACy. β€’ Experience with Generative AI and Large Language Models (LLM) β€’ Evidence of true self-starter and operating independently. β€’ Fluency in Python Programming, version control and collaboration with GIT, standard Python packages (ex. Pandas, numpy, matplotlib) and ML frameworks β€’ Knowledge of TensorFlow, PyTorch, Pandas, scikit-learn, NLTK, Azure ML (optional), Amazon Web Services EC2. β€’ Experience with scalable data engineering frameworks such as Apache Spark and orchestration frameworks such as Airflow, and/or experience with semantic search. β€’ Expert knowledge in conducting data analysis and applying advanced statistical concepts and ML methods to build, train, test, and evaluate a variety of supervised and unsupervised analytic models. β€’ Experience with ML model deployment and operations like DevOps, MLOps, LLMOps. β€’ Experience with NLP and Generative AI libraries like regular expressions (e.g., spacy, langchain), text annotation tools and semantic frameworks. β€’ Ability to clean and process large amounts of real-world data. β€’ Experience retrieving and manipulating data from a variety of data sources included DB2, Oracle, SQL Server, Hadoop and flat files. β€’ Excellent Communication skills. β€’ Experience with database management systems (e.g., PostgresSQL, MySQL, SQLite, SQL, etc.) β€’ Excellent analytical skills to identify potential risks and propose effective solutions. β€’ Excellent problem-solving skills, ability to collaborate with cross-functional teams and proven communication in written and verbal formats to various audiences to include executive leadership. β€’ Prior experience with federal or state governments IT projects. β€’ Industry experience preferred β€’ Experience with, or the ability and willingness to learn distributed processing via the Hadoop ecosystem, i.e., Spark, Impala and Hive. β€’ Experience working in an analytical research environment. β€’ Experience in parallel processing such as GPU programming with CUDA β€’ Experience with Mathematica β€’ Experience using markup languages such as LaTeX, HTML, etc. β€’ Experience with Natural Language Processing for anomaly detection.