TV

t victor

Contact details — on requestProof of Work — on request

Experience

C

Data Engineer

Covalense Global Pvt Ltd · Oct 2025 – Apr 2026

Not yet confirmed
  • Worked on an existing Enterprise Data Lake Analytics and Storage Optimization project using Azure Databricks, PySpark, SQL, and ADLS Gen2 to support enterprise reporting and storage analysis.
  • Developed and supported PySpark-based ETL workflows to process weekly Azure Blob Inventory data stored in ADLS Gen2 and prepare curated datasets for business reporting.
  • Applied data transformations, filtering, joins, aggregations, validation, and reconciliation to generate clean and reliable datasets for analytics.
  • Calculated storage metrics, analyzed storage utilization trends, identified inactive data paths, and prepared reporting datasets to support storage optimization initiatives.
  • Used SQL for data validation, reconciliation, and analysis to ensure data accuracy before publishing datasets for reporting.
  • Prepared curated datasets consumed through Azure Synapse Analytics and Power BI dashboards for storage analytics and business reporting.
  • Supported production activities by monitoring Databricks jobs, troubleshooting failures, rerunning failed jobs when required, validating processed data, and resolving production issues.
  • Monitored workflow executions using Azure Monitor, analyzed execution logs, and supported issue resolution to maintain reliable data processing.
  • Supported CI/CD deployment activities and version control using Azure DevOps and Git for Databricks notebooks and deployment changes.
  • Worked with Unity Catalog and RBAC concepts to support secure, governed, and controlled access to enterprise datasets.
  • Collaborated with business users, reporting teams, and project stakeholders in an Agile environment to deliver enhancements and support reporting requirements.
O

Data Engineer

Optimum Solutions Pvt Ltd · Oct 2024 – Apr 2025

Not yet confirmed
  • Worked on an existing Transaction Monitoring and Analytics Platform using Azure Data Factory, Azure Databricks, PySpark, SQL, and Azure Event Hub to support batch and real time transaction processing for fraud monitoring and compliance reporting.
  • Developed PySpark transformation and cleansing logic in Azure Databricks following Medallion Architecture (Bronze, Silver, Gold) to process digital wallet transaction data and prepare curated datasets.
  • Implemented fraud detection validations such as duplicate transaction checks, velocity checks, unusual transaction patterns, and failed login monitoring to identify suspicious user activities.
  • Processed and stored curated transaction data in Delta Lake using incremental loading, schema evolution, and time travel features to support reliable and scalable data processing.
  • Supported production pipelines by monitoring job executions, troubleshooting failures, rerunning failed jobs when required, performing basic Root Cause Analysis (RCA), and resolving production issues.
  • Monitored pipeline executions using Azure Monitor and Log Analytics, analyzed execution logs, and supported issue resolution to maintain reliable pipeline performance.
  • Improved pipeline performance by optimizing PySpark transformations, SQL queries, partitioning strategies, and Delta Lake operations to reduce processing time.
  • Supported CI/CD deployment activities and version control using Azure DevOps and Git for Azure Data Factory pipelines, Databricks notebooks, and deployment changes.
  • Worked with RBAC, Azure Key Vault, and Managed Identities to support secure access and protection of financial and customer data within the Azure environment.
S

Data Engineer

SLK Software · Jul 2020 – Sep 2024

Not yet confirmed
  • Developed Azure Data Factory pipelines to ingest customer data from ERP, CRM, SQL Server, and other enterprise systems into ADLS Gen2.
  • Created ADF datasets, linked services, activities, and pipeline workflows to support scheduled and incremental data processing.
  • Developed PySpark notebooks in Azure Databricks to clean, transform, join, and aggregate customer data.
  • Applied filtering, joins, aggregations, deduplication, and business rules during data transformation.
  • Used SQL to validate source and target data, perform reconciliation, and investigate data-related issues.
  • Implemented data quality checks for null values, duplicate records, record counts, and required fields before publishing datasets.
  • Prepared curated datasets in ADLS Gen2 for downstream analytics and reporting requirements.
  • Processed data through different stages of the data pipeline and followed structured data processing practices for reliable downstream consumption.
  • Optimized PySpark transformations, SQL queries, and ADF pipeline activities to improve processing performance.
  • Monitored scheduled ADF pipelines and Databricks jobs and investigated failures during daily production activities.
  • Prepared analytics-ready datasets for Azure Synapse Analytics and Power BI reporting.
  • Supported deployments, code changes, and version control using Azure DevOps and Git while working in an Agile environment.
  • Designed, developed, and supported real-time ETL pipelines and streaming workflows to ingest sales transactions from POS systems, e-commerce platforms, and warehouse applications using Kafka and Azure Stream Analytics.
  • Developed PySpark-based transformation logic in Azure Databricks to process streaming and batch datasets, perform data cleansing, filtering, joins, aggregations, and prepare curated datasets.
  • Processed streaming data with Azure Stream Analytics and loaded transformed datasets into Azure Synapse Analytics for analytical reporting and business intelligence.
  • Developed and optimized SQL queries to validate streaming data, perform reconciliation, and support reporting.
  • Built analytical data models and prepared reporting datasets for Power BI dashboards to monitor sales performance, revenue trends, and operational KPIs.
  • Implemented incremental data processing, data validation, reconciliation, and exception handling to ensure data accuracy and consistency.
  • Supported production pipelines by monitoring job executions, troubleshooting failures, resolving defects and performing basic Root Cause Analysis (RCA).
  • Improved pipeline performance by optimizing PySpark transformations, SQL queries, partitioning strategies, and streaming data processing.
  • Supported CI/CD deployments and version control using Azure DevOps and Git Applied RBAC and data masking techniques to provide secure and controlled access to business data.
  • Collaborated with business stakeholders and reporting teams to gather requirements and deliver scalable real-time analytics solutions.

Skills 0 proven through work

Also works with

Business Key IdentificationData OptimizationInactive Storage IdentificationStorage Growth CalculationStorage Usage AnalysisData DeconsolidationData RemovalCode ReviewLog AnalysisStakeholder Communication AdaptationProcess RerunningJupyter NotebookWork Planning and PrioritizationRequirements AnalysisAgile MethodologyScalable System DesignOn-time DeliveryTeamworkTechnical Skill DevelopmentEnd-to-End Data Science Lifecycle ManagementData EngineeringData Integrity ManagementData Unification and NormalizationData PreprocessingERP Software ManagementCRM ManagementData CollectionPython DevelopmentBulk Data ProcessingInput Validation ImplementationData Repetition PreventionDuplicate Payment Prevention ImplementationFraud DetectionData MonitoringProduction Environment SupportData Discrepancy InvestigationMedallion Architecture DesignTabular Data Cleaning & ValidationAzure Data Factory Pipeline DevelopmentReal-Time Data IntegrationDatabase Query OptimizationFeature SelectionAzure Blob Storage IntegrationCustom Metric DesignComparative Performance AnalysisQlik Sense UsageSSMS UsageAPI-Based Data CollectionData CatalogingVersion Control with GitHubAzure DevOpsApplication DeploymentTechnical TroubleshootingETL Pipeline DevelopmentPower BIAzure Synapse Analytics UsageAnalytical Report CreationData ValidationData TransformationPySparkAzure Data Lake UsageQuery WritingSpark SQLDatabricks Usage

Proof of Work

Proof of Work

t shares this with people who ask. You'll hear back either way.

Education

Bachelor's Degree, Bachelor of Technology

Jawaharlal Nehru Technological University Anantapur

Contact details

Contact details

t shares this with people who ask. You'll hear back either way.

Ask about this candidate

AI responses may contain errors. Proof status and source information are provided by Proof of Skill.

A Proof CV — one profile, kept current, with the proof attached. Back to the portal