How To Set Up Zeppelin For Analytics And Visualization
In this article, you learn how to create and configure a Zeppelin instance on an EC2, and about notebook storage on S3, and SSH access.
...technologies may include: DevOps/Cloud/SRE: AWS, Azure, GCP, Kubernetes, Docker, Terraform, CI/CD, Linux, monitoring, infrastructure automation and DevSecOps. AI/ML: Python, PyTorch, TensorFlow, scikit-learn, generative AI, LLMs, NLP, computer vision, MLOps, model deployment and cloud ML platforms. Data Engineering: Python, SQL, Spark, Databricks, Snowflake, BigQuery, Redshift, Airflow, Kafka, ETL/ELT, data lakes and data warehouses. How to Apply: Send a direct message containing: • Your résumé • LinkedIn profile • Primary technical area • Years of relevant experience • Current availability • A short pre-interview video introducing yourself, describing your technical background and explaining one recent project ...
...Candidates without formal certification are encouraged to apply if they have substantial experience developing with the Zoho ecosystem. Relevant experience includes: Zoho Recruit API Zoho CRM API Zoho Books API Zoho OAuth/API authentication Zoho custom modules and fields Zoho record relationships Zoho attachments/files Deluge Supabase PostgreSQL SQL REST APIs Backend development Database migrations ETL/data pipelines Authentication and authorization Production system maintenance Experience with ATS platforms, recruitment systems, HR software, or job portals would also be helpful. When Applying Please include: Your experience working with Zoho Recruit, Zoho CRM, or Zoho One Examples of projects where you worked directly with Zoho APIs or Deluge Any Zoho certifications or Par...
...learning classification models to enhance anomalous behavior detection. • Achieved 88% accuracy in identifying high-risk taxpayers through advanced predictive modeling. • Developed a scalable analytics pipeline using PySpark for distributed data processing and an interactive Streamlit dashboard with 7 analytical views with KPIs, Tables and Graphs for real-time compliance monitoring. Built an end-to-end ETL data analytics pipeline with 500K+ transaction records for customer purchasing insights. • Preprocessed the dataset, performed an 80/20 train-test split, and conducted EDA and feature engineering. • Designed customer segmentation ML models using RFM analysis to classify users into 3 behavioral groups. • Improved decision-making and strategic planning by...
...NumPy. Made an automated ETL/data preprocessing pipeline to take care of data cleaning, changing data making it all the same and checking if it is all the same. Added a PostgreSQL (Neon) database to keep data in a good way and help with machine learning work. Did analysis of the data. Looked at the features to find what helps predict salary. Trained a Random Forest regression model to guess salaries. Found that employee_residence is one of the things to use for guessing salaries. Built a FastAPI backend with places to get analytics and predictions. Linked the machine learning backend to a React frontend to show information and results in a way. Technologies: Python, Pandas, NumPy, Scikit-learn, Random Forest, PostgreSQL, Neon, FastAPI, React, ETL. This...
I need an exp...SQL database after applying several transformations. The raw worksheet contains multiple sheets of mixed data types, so before anything is loaded I’ll need column mapping, type casting, and a few calculated fields added (we can review the exact logic together). Once the transformation rules are approved, please import everything into the live SQL instance and provide: • A repeatable script or ETL package (Python, SSIS, or similar) • A short README explaining how to rerun the job and tweak the mappings • Verification that every record landed in the correct table with accurate data types and relationships I’ll supply the sample XLS, the current SQL schema, and the transformation notes as soon as we start. Looking forward to a smoo...
...can also grant read-only access to the underlying SQL database if a live connection is preferable. Once the data model is clean, I’d like a series of drill-through dashboards that let stakeholders slice performance by region, product line, and individual rep, complete with trend indicators and shareable links. Deliverables • A tidy, well-documented data model (joins, calculated fields, and any ETL scripts) • One core “Efficiency Snapshot” dashboard plus supporting views for deeper dives • A short video or live walkthrough so the team can maintain and extend the visuals I’ll handle deployment; you focus on insight-rich design, clear labeling, and responsive performance. Let me know which tool you’ll lean on first and how quick...
AI-Powered Sales & Operations Reporting System with Automated ETL Pipeline and Power BI Dashboard We are looking for an experienced Data Analytics & Automation Developer to build a fully automated reporting solution that consolidates business data from multiple sources into a centralized Power BI dashboard. Project Overview: Our company currently relies on several disconnected systems including Excel reports, SQL databases, CRM exports, and online APIs. Reporting is manual, time-consuming, and prone to errors. We require an automated data pipeline that extracts, cleans, validates, and combines data into a single reporting database while providing executive-level Power BI dashboards with real-time KPIs. The selected freelancer will be responsible for designing and implemen...
AI-Powered Sales & Operations Reporting System with Automated ETL Pipeline and Power BI Dashboard We are looking for an experienced Data Analytics & Automation Developer to build a fully automated reporting solution that consolidates business data from multiple sources into a centralized Power BI dashboard. Project Overview: Our company currently relies on several disconnected systems including Excel reports, SQL databases, CRM exports, and online APIs. Reporting is manual, time-consuming, and prone to errors. We require an automated data pipeline that extracts, cleans, validates, and combines data into a single reporting database while providing executive-level Power BI dashboards with real-time KPIs. The selected freelancer will be responsible for designing and implemen...
...Cloud (IaaS) deployed system: • Work in Test environment, think ahead for Production • Document env vars + deployment requirements • Coordinate with Infomaniak (domain, DB-URL, storage, logs) • Use Linux/SSH basics (not full sysadmin) Nice to Have • FinTech / WealthTech / Stock-Market experience • Complex responsive tables • Multi-step wizards • Payment integrations • Data pipelines, ETL, crawling, async API calls • AI-assisted coding by (architectural rules, naming conventions) Your Profile • English mandatory (code + documentation) • Senior-level full-stack developer • Strong Python + strong JavaScript/TypeScript • Experience with Django REST + Svelte 5 • Skilled in parsing, data engi...
...Cloud (IaaS) deployed system: • Work in Test environment, think ahead for Production • Document env vars + deployment requirements • Coordinate with Infomaniak (domain, DB-URL, storage, logs) • Use Linux/SSH basics (not full sysadmin) Nice to Have • FinTech / WealthTech / Stock-Market experience • Complex responsive tables • Multi-step wizards • Payment integrations • Data pipelines, ETL, crawling, async API calls • AI-assisted coding by (architectural rules, naming conventions) Your Profile • English mandatory (code + documentation) • Senior-level full-stack developer • Strong Python + strong JavaScript/TypeScript • Experience with Django REST + Svelte 5 • Skilled in parsing, data engi...
...it, then runs machine-learning routines so the results can be consumed by dashboards, reports, or downstream models. The exact data source—stocks, crypto, or transactional feeds—is still being finalised, so the solution must stay modular enough to swap connectors without large rewrites. I expect the work to centre on Python with pandas, NumPy, scikit-learn (or similar), and a well-structured ETL workflow orchestrated by notebooks or a lightweight API layer. Good documentation and clean, reproducible code are essential; once delivered, my in-house team must be able to extend the models or plug in new data streams without your help. Deliverables • A fully functioning preprocessing and feature-engineering module (cleansing, deduplication, outlier handling, e...
I have a raw dump of sales data that I want turned into a clear, insight-driv...I expect: • A well-structured, self-updating workbook with labeled sheets for raw data, cleaned data, and final dashboards • Interactive visuals (line or column charts, slicers, conditional formatting) that highlight key sales trends, peaks, and dips • A short written summary within the workbook explaining the main findings and formulas used If you prefer Power Query or Power Pivot for the ETL step, feel free—just keep everything contained in the .xlsx file so it runs on a standard desktop install of Excel 2019 or later. Accuracy and clarity matter more to me than fancy design; by the end I want to open the file, change or append new sales rows, hit refresh, and immediately see ...
...be a small ETL project. I want to build a realistic, enterprise-level data platform that demonstrates how the different components of a modern data engineering ecosystem work together. The project should cover the complete data lifecycle: Data Sources → Ingestion → Streaming → Batch Processing → Data Lake → Transformation → Data Warehouse → Data Quality → Orchestration → Analytics → Monitoring → Deployment I also want the freelancer to explain the architecture and technologies clearly throughout the project so that I can understand why each technology is being used, what problem it solves, and how the components integrate with each other. Technical Stack: Python, SQL, Apache Kafka, Spark/PySpark, Databricks, Delta Lake, A...
... Proven record of automating client workflows, building self-serve Streamlit apps, scaling dbt pipelines to 600+ clients, and improving query performance by up to 95%. Expert in RBAC, data governance, CI/CD, and Snowpark-driven transformations across AWS and Azure. CORE TECHNICAL SKILLS Cloud /Dataplatforms Snowflake, dbt (Core, Cloud), SnowSQL, Snowpark, Databricks (Lakehouse), AWS, Azure ETL / Orchestration Programming BI & Applications Apache Airflow, Fivetran, Prefect, Snowpipe, Streams & Tasks Python, SQL, PySpark, Stored Procedures, UDFs, Snowpark Python, Pandas, NumPy Tableau, Streamlit (Snowflake Native Apps), Power BI DevOps & CI/CD Governance & Security Databases Git, GitLab CI/CD, GitHub Actions, Agile, Query Optimization, Performance Tuni...
I run a commercial real-estate portfolio on Yardi Voyager 8 and need an expert who knows the platform inside out. The immediate focus is a full migration of tenant information, financial data, and property details from legacy spreadsheets and third-party systems into our live Voyager 8 database. Data integrity and audit trails are non-negotiable, so familiarity with Spreadsheet Import, ETL tools, SQL, and YSR validation reports will be essential. Once the data is in place, I want to reshape several standard reports so they speak our team’s language—cash-flow snapshots, rent-roll drill-downs, and KPI dashboards that can be scheduled or exported to Excel without manual cleanup. If you have experience tailoring Crystal or SSRS reports for Voyager, please highlight it. Fin...
...datasets Develop data transformations using PySpark and SQL Work with Delta Lake and Databricks Lakehouse architecture Troubleshoot pipeline and performance issues Improve reliability, scalability, and performance of existing workloads Collaborate with our engineering team on technical requirements Required Skills Strong hands-on Databricks experience PySpark / Apache Spark Python SQL Delta Lake ETL/ELT and data pipeline development Data modeling Git and CI/CD experience Experience with cloud-based data platforms Nice to Have Databricks Workflows / Jobs Unity Catalog Structured Streaming Performance optimization AWS, Azure, or GCP experience Databricks certification What We're Looking For We need someone who has real production experience with Databricks, not only theoretic...
...business requirements into technical requirements and actionable user stories. • Lead initiatives from discovery → implementation → measurement → optimization. Technology Exposure • CRM: Braze, Salesforce Marketing Cloud, Iterable, CleverTap, MoEngage • CDP: mParticle, Segment • Data: Azure, Snowflake, BigQuery, Redshift • Analytics: GA4, Amplitude, Looker, Power BI • Activation: Census / Reverse ETL • Tracking: Google Tag Manager, UTM & event tracking • Channels: Email, SMS, Push, WhatsApp What We’re Looking For • 6+ years in CRM, Lifecycle Marketing, Marketing Automation or MarTech. • Hands-on experience with an enterprise CRM platform. • Strong understanding of lifecycle, segmentation, personaliz...
...possible, full names—without sacrificing accuracy. Here’s what I need: • A clear proposal for a new enrichment workflow or alternative provider that keeps per-record costs low while maintaining dependable match rates. • A small pilot run (10–50 k records) so I can verify accuracy, error handling, and throughput before we scale. • A documented pipeline—scripted in Python, R, or a lightweight ETL tool of your choice—showing how to push millions of records through the chosen API/service, manage rate limits, and export the enriched data in CSV or Parquet. • A concise comparison sheet outlining projected monthly costs versus Melissa and any licensing caveats. Acceptance criteria: the pilot must achieve match rates comparable t...
... Stack for Phase 1: GitHub + Codespaces (free allowance) + Python + SQLite + Streamlit + OpenAI API + Scheduler + GitHub Secrets. Using Codex as agentic engineer. What I will build: 1. Ingest: Weather (Open-Meteo/NOAA), Gov Reports (USDA WASDE/FAO PDF scraper + LLM summary), News RSS, Prices (Yahoo/FRED). Modular for satellite/NDVI later. 2. Core: SQLite schema (raw_data, evidence, scores), ETL, Scoring 0-100 [30% Supply Stress, 25% Demand, 20% News Sentiment, 15% Price Anomaly, 10% Risk], Anomaly detection. 3. Challenge/Risk-check: 2+ source cross-check, contradiction check, low-confidence flag if evidence <3, audit log. 4. Present: Streamlit dashboard—Ranked watchlist (Commodity, Score, Trend, Risk, Evidence Count, Last Signal), Evidence drawer with URLs and filt...
...Stack for Phase 1: GitHub + Codespaces (free allowance) + Python + SQLite + Streamlit + OpenAI API + Scheduler + GitHub Secrets. Using Codex as an agentic engineer. What I will build: 1. Ingest: Weather (Open-Meteo/NOAA), Gov Reports (USDA WASDE/FAO PDF scraper + LLM summary), News RSS, Prices (Yahoo/FRED). Modular for satellite/NDVI later. 2. Core: SQLite schema (raw_data, evidence, scores), ETL, scoring 0-100 [30% supply stress, 25% demand, 20% news sentiment, 15% price anomaly, 10% risk], anomaly detection. 3. Challenge/Risk-check: 2+ source cross-check, contradiction check, low-confidence flag if evidence <3, audit log. 4. Present: Streamlit dashboard—Ranked watchlist (Commodity, Score, Trend, Risk, Evidence Count, Last Signal), Evidence drawer with URLs and fil...
...and keep delivery on track through solid project-management routines. Because most of our insights live in raw data, I need someone comfortable jumping into Python scripts for quick analyses or automation, then pivoting straight into Excel for deeper financial or operational modeling. Must-have toolbox: • Data analysis, Project management, Stakeholder communication • Python for ad-hoc queries/ETL tasks • Excel for advanced modelling and clear visual reports If you pair meticulous documentation with the ability to speak the language of both C-suite leaders and software engineers, let’s talk. Share a brief note on a recent tech project where your analysis directly influenced product direction, along with any artifacts (user stories, process maps, dashbo...
I am making a fast-track move into data engineering and need an experienced tutor who can work with me right away. Our first deep-dive will centre on Apache Iceberg; once I am comfortable there we can branch into Spark or Kafka if that helps reinforce the concepts. The practical objective is to design and build production-ready ETL pipelines from ingestion through transformation and loading. I will be practicing on both AWS and Azure, so please be comfortable switching between the two clouds and showing me the platform-specific nuances of storage, compute and orchestration services. Sessions should be highly hands-on—screen sharing, live coding, short homework labs and code reviews rather than slide decks. My target is to complete the learning path progressively by the end o...
...ground up and want an experienced AWS Data Engineer to act as a personal tutor/mentor — covering core concepts, tools, and hands-on skills needed to work as a data engineer on AWS. What I'm looking for: 1:1 sessions (video call) covering AWS data engineering fundamentals and hands-on labs Coverage of core services: S3, Glue, Redshift, Athena, Kinesis, Lambda, EMR, Step Functions, RDS/DynamoDB ETL/ELT pipeline design, data lake vs. data warehouse concepts Python/PySpark for data engineering, SQL for analytics Guidance on real-world project work I can add to my portfolio/resume Optional: help prepping for AWS Data Engineer – Associate certification Ideal candidate: Proven hands-on experience as an AWS Data Engineer (production experience preferred, not just cer...
I’m putting...chart, and drill-through has been built with performance analytics in mind, so you can spot slow pages, high-exit funnels, and untapped revenue opportunities without touching code. You will receive: • Full, well-documented Python source code • Ready-to-import SQL database (schema + seed data) • Three PowerBI dashboards (overview, deep-dive, executive snapshot) Import the database, point the ETL to your own logs, and the visuals light up—no extra licensing or add-ons required beyond the standard PowerBI desktop/service. The price is negotiable and I’m keen to work only with buyers who are ready to move quickly. Let me know any specific KPIs or custom filters you’d like and I’ll confirm they’re already covered o...
...Preferred Skills: - Strong experience in Python backend development, preferably FastAPI, Flask, or Django - Experience with data engineering workflows - Knowledge of databases such as PostgreSQL, MySQL, or MongoDB - Experience with REST APIs - Basic understanding of React frontend integration - Ability to write clean, maintainable, and well-documented code Nice to Have: - Experience with Pandas, ETL pipelines, background jobs, or cloud storage - Experience deploying backend services - Understanding of authentication and role-based access We are looking for someone who can understand the requirement, suggest the right backend architecture, and build the first working version quickly. Please share relevant backend or data engineering project examples, your suggested tech stack,...
...without developer help. • Data visualization tools that present trends clearly for non-technical decision-makers. I’m open to the platform you recommend—whether that’s Snowflake, SQL Server, BigQuery, Redshift, or a comparable HIPAA-compliant stack—as long as you outline why it fits our scale and security needs. Please plan for typical healthcare data standards (HL7 v2, FHIR, or CCD) and include an ETL/ELT pipeline that cleans, normalizes, and de-identifies data where required for compliance. Deliverables expected: 1. Dimensional warehouse schema and data-dictionary documentation. 2. Automated ingestion scripts/workflows and orchestration setup. 3. Interactive dashboards plus a library of reusable report templates. 4. Deployment guide and ...
...application * Build data-quality checks, anomaly detection and confidence systems * Document data sources, transformations, assumptions and limitations * Work closely with the founder and development team to translate product requirements into reliable data infrastructure Technical skills we are looking for Strong experience with: * Python * SQL * PostgreSQL * APIs / REST APIs * Web scraping * ETL / ELT pipelines * Data cleaning and normalization * Pandas / NumPy * Large structured datasets * Automation * Scheduled data pipelines * Geospatial data and coordinate matching * Git / GitHub Experience with some of the following is a major advantage: * Machine learning * Recommendation engines * Ranking and scoring algorithms * Time-series data * Marine weather data * Oceanographi...
...Data Engineer with 5+ years of experience in designing, developing, and implementing data migration and archival solutions. The ideal candidate will have expertise in Azure cloud technologies, ETL/ELT development, data validation, and enterprise data migration projects involving ERP and HCM systems. This role will focus on building scalable data pipelines, ensuring data quality, and supporting large-scale migration and archival initiatives. Technologies: Azure Data Lake, Azure Data Bricks, Azure Data Factory, Azure DevOps, Gitlab, Github, & Snow Key Responsibilities: • Design and implement ETL processes to move data from various legacy systems to Azure data platforms. • Design and develop data orchestration workflows using Azure Data Factory (ADF)...
...pipeline. The system will ingest user onboarding files, execute real-time credit and fraud verification checks via third-party financial APIs, process unstructured financial statements utilizing LLMs, and route qualified profiles into an executive Power BI KPI dashboard while instantly alerting compliance officers via Slack and CRM updates. Project Requirements: - Intelligent Data Ingestion & ETL Pipelines: Build automated web scraping and data extraction pipelines to aggregate market compliance data and transactional parameters, while setting up secure webhook listeners and API connectors to ingest client onboarding forms and financial documents instantly. - AI-Powered Risk Analysis & Decision Engine: Integrate advanced Large Language Models via API (OpenAI, Claude) ...
...Google Data Studio are all acceptable as long as the final product is fast, visually clear, and can connect directly to my SQL schema. I’ll provide database credentials and an ERD. You’ll handle the data pull, modeling, and visual layout, then publish or package the dashboard so I can share it company-wide. Deliverables: • Live dashboard connected to the SQL source • Clean, commented SQL (or ETL) queries used for the model • Brief hand-off video or PDF walkthrough covering filters, refresh schedule, and how to extend the visuals Acceptance criteria: 1. Monthly revenue growth displayed as both numeric and trend line. 2. Filters for time range and region respond in under two seconds on a standard laptop. 3. All calculations reproduce the figures...
...away. The focus is pure data integration: designing and orchestrating Fabric pipelines, unifying data from disparate sources, and getting everything landed, cleaned, and ready for downstream consumers. While some reporting or admin work may pop up, the core responsibility is building robust, production-grade integration workflows inside Fabric itself—knowledge of only Power BI, SSIS, or generic ETL will not be enough. The engagement starts at a guaranteed 20 hours a week, with the real possibility of expanding to a full 40 as we move deeper into the roadmap. Work from home is perfectly fine, and for compliance and collaboration reasons I’m limiting the hire to India-based professionals. To be considered, send a brief but detailed project proposal highlighting: &bull...
...up and tune clusters, build out end-to-end ETL pipelines, and craft scalable, repeatable machine-learning jobs that flow smoothly into production. Beyond the build, I rely on you to keep the lights on. Our platform runs 24/7, so swift, senior-level ticket resolution is critical—whether it’s a misbehaving notebook, a flaky job, or a performance bottleneck inside a Spark cluster. You should be perfectly comfortable navigating Azure services, AKS/Kubernetes, Docker, and Python/PySpark, bringing best-practice CI/CD and full MLOps rigor to the table. Key deliverables • Cluster setup and performance tuning in Azure Databricks • Production-ready ML workflows with versioned models, automated tests, and monitoring hooks • Robust data-integration/ETL...
Modify Databricks PySpark ETL Notebook with Dimension Joins and Aggregation Changes --- ## **Project Title** Modify Databricks PySpark ETL Notebook with Dimension Joins and Aggregation Changes --- ## **Project Description** I have an existing Databricks PySpark ETL notebook that creates a Delta table from a source fact table. I need to modify the transformation logic based on updated business requirements. ### Existing Environment * Databricks (Serverless) * PySpark * Unity Catalog * Delta Lake * Existing notebook uses metadata framework, DQ framework, AddAudits(), ValidateSchema(), DynamicOverwrite(), etc. --- ### New Requirements Need to enrich the fact data using three dimension tables. ### 1. Business Mapping Join: ``` fact.material_number =
Summary Need a Data Pipeline Engineer (Python / ETL / PostgreSQL) We are looking for an experienced Data Pipeline Engineer to integrate multiple food databases into an existing production backend. Important: This is not a greenfield project. The backend, database schema, APIs, and food intelligence engine have already been built. Your role is to build and maintain the ingestion pipelines that feed the existing system. Current Backend Status * Existing PostgreSQL database schema * Existing backend APIs * Existing food scoring engine * Existing product model and normalization framework * Existing infrastructure for product ingestion Scope of Work Build reliable ingestion pipelines to import, clean, normalize, validate, and synchronize data from the following sources: * Open Foo...
...SQL Server Integration Services (SSIS), and SQL Server Reporting Services (SSRS). The candidate will be responsible for designing and developing ETL/ELT pipelines, creating and optimizing SQL queries, developing reports and dashboards, managing data integration processes, ensuring data quality, and supporting business intelligence and data warehousing initiatives. The ideal candidate should have strong analytical and problem-solving skills, experience with database performance tuning, and the ability to work collaboratively in a remote team environment. Prefered Skills Microsoft SQL Server, Azure Data Factory (ADF), Azure Databricks, SSIS, SSRS, ETL Benefits: Work from home Application Question(s): Are you comfortable working during the shift from 9:00 PM to 1:00 AM IS...
Brief B: Field Sales Mapping & Route Optimisation Tool Job title: Build an Interactive ...Directions API with waypoint optimisation. Backend algorithm for TSP/greedy subset selection. Ability to connect to external REST APIs. Experience with field sales or logistics tools is a plus. Budget & Timeline Timeline: 4–5 weeks. How to apply (both briefs) Please include: Your understanding of the project in your own words. Links to 1–2 similar projects (especially with map + routing for Brief B, or ETL/scoring for Brief A). Your proposed tech stack. Estimated hours and total cost. Earliest start date. I’m ready to start immediately with the right candidate. If you can build both parts (intake + mapping), please mention that – I may award both to one person/te...
...cloud VM. Code quality: Environment variables for secrets, clear README, minimal test coverage. Deliverables Fully working backend with admin UI (can be minimal, functional). API documentation (OpenAPI/Swagger or simple markdown). Sample CSV and test outcomes that prove the scoring and API flow. Setup guide. Skills needed Backend development, experience with CSV ETL pipelines. Google Geocoding API. Basic scoring logic implementation. Creating clean APIs. (Plus) Experience with lead scoring or construction data. Budget & Timeline...
Professional Summary Results-driven Azure Data Engineer and Power BI Developer with 13+ years of experience in Business Intelligence, Data Warehousing, Azure Data Platform, Microsoft Fabric, Power BI, SQL Server, SSIS, and Azure Synapse Analytics. Experienced in designing scalable ETL pipelines, developing interactive Power BI dashboards, migrating on-premises solutions to Azure, and delivering enterprise reporting solutions. Strong expertise in SQL performance tuning, Azure Data Factory, Synapse Pipelines, Microsoft Fabric, Git, Azure DevOps, and data modeling. Adept at working with global clients in Agile environments and delivering freelance consulting assignments independently. Key Skills • Cloud Platforms: Azure Data Factory, Azure Synapse Analytics, Azure SQL, Microsoft F...
...2–3 channels you would prioritize for AutoVel Extract in the first 30 days, and why. D. A short 30-day plan (week by week). E. Your rate / retainer expectation and weekly availability. Start your proposal with this exact line: I explored AutoVel Extract and my first ICP is: … Proposals missing that line will be skipped. Nice to have Experience marketing tools for dealers, ecommerce ops, data/ETL, scraping-to-structured-data, or B2B productivity SaaS Basic CRO / landing page testing experience Comfortable working async with a technical founder...
...cleanly into Microsoft Project. The scope is limited to resource information only—tasks and schedules are handled elsewhere—so the work revolves around mapping each resource, its role, and any availability or workload attributes we decide are essential once we talk through the final field list. You may choose the tooling you prefer (PowerShell, VBA, Project’s built-in import wizard, or a lightweight ETL script) as long as the end result is an MS Project file where every resource from the CSVs appears correctly and is ready for further scheduling. I’ll provide a sample CSV and the full data set after kickoff; you’ll return a working .MPP (or a reproducible script that generates it) plus brief setup notes so I can repeat the process later. Acceptance...
Project Description I'm looking for an experienced Azure Databricks / PySpark developer to help convert an existing SQL Server view into a Databricks PySpark transformation. Project Overview I have: An existing SQL view with multiple CTEs. An existing Databricks Asset ...PySpark implementation. Clean, optimized, production-ready code. Logic matching the SQL view. Compatible with Databricks Runtime. Delta table creation. SQL View creation. Assistance with testing if required. Required Skills Azure Databricks PySpark Spark SQL Delta Lake SQL Server DataFrame API Window Functions Databricks Asset Bundles (preferred) Nice to Have Experience with Unity Catalog Production ETL development Azure Data Factory knowledge What I'll Provide Existing SQL View Existing Databricks not...
...two-way sync that moves customer, sales, and product data automatically—no more CSV exports or manual double-entry. Here is how I picture the job flowing: • Map every required Salesforce object (Accounts, Contacts, Opportunities, Products, Orders) to the corresponding tables in the ERP and establish field-level rules for creates, updates, and deletes. • Build the middleware or use a proven iPaaS/ETL framework—whichever achieves near-real-time transfers with solid error handling and retry logic. • Expose inventory, pricing, and order status from the ERP back into Salesforce so the sales team always sees live figures. • Preserve the ERP’s accounting integrity; journal entries and financial reports must remain balanced after each sync cy...
I have an in-house project that needs a fully custom-built data pipeline. The goal is to ingest Parquet files arriving in our file system, apply the required transformations, and move the cleaned data on to our analytics environment. Off-the-shelf ETL tools do not meet our requirements, so I am looking for a bespoke solution designed from the ground up. Scope of work • Design the end-to-end pipeline architecture, choosing the most suitable framework (for example Python, Spark, or similar) while keeping future scalability in mind. • Build robust code that can automatically detect new Parquet files, validate their schema, transform the data as specified, and deliver the output to the target location I will provide once we start. • Add logging, error handling, and...
...SQL Server Integration Services (SSIS), and SQL Server Reporting Services (SSRS). The candidate will be responsible for designing and developing ETL/ELT pipelines, creating and optimizing SQL queries, developing reports and dashboards, managing data integration processes, ensuring data quality, and supporting business intelligence and data warehousing initiatives. The ideal candidate should have strong analytical and problem-solving skills, experience with database performance tuning, and the ability to work collaboratively in a remote team environment. Prefered Skills Microsoft SQL Server, Azure Data Factory (ADF), Azure Databricks, SSIS, SSRS, ETL Benefits: Work from home Application Question(s): Are you comfortable working during the shift from 9:00 PM to 1:00 AM IS...
I have an in-house project that needs a fully custom-built data pipeline. The goal is to ingest Parquet files arriving in our file system, apply the required transformations, and move the cleaned data on to our analytics environment. Off-the-shelf ETL tools do not meet our requirements, so I am looking for a bespoke solution designed from the ground up. Scope of work • Design the end-to-end pipeline architecture, choosing the most suitable framework (for example Python, Spark, or similar) while keeping future scalability in mind. • Build robust code that can automatically detect new Parquet files, validate their schema, transform the data as specified, and deliver the output to the target location I will provide once we start. • Add logging, error handling, and...
...Automation - Playwright (Mandatory) - Automated UI testing - End-to-End (E2E) Testing - Regression Testing - Functional Testing - Integration Testing 2. API Testing - REST APIs - API Validation - Request/Response Verification - Authentication Testing - Error Handling 3. Performance Testing - JMeter (Preferred) - Load Testing - Stress Testing - Performance Analysis 4. Data Engineering Testing - ETL Testing - ELT Testing - Batch Data Validation - Real-time Data Validation - Streaming Pipeline Testing - Data Transformation Validation - Data Quality Validation 5. Database & SQL - Strong SQL - Data Validation - Data Reconciliation - Data Warehouses - Database Testing 6. Reporting & Analytics - Tableau Dashboard Testin...
...emerging trends, and highlight performance bottlenecks whenever those needs arise. You are free to recommend the stack—whether that means a lightweight Python pipeline with pandas and SQLAlchemy, a BI layer in Power BI or Tableau, or a full-blown cloud-native approach—so long as it integrates cleanly with our existing database schema and minimal overhead for maintenance. Key deliverables • Clean ETL/ELT workflow that connects to the internal database, performs any required transformations, and keeps data fresh on an agreed schedule • Metric definitions documented in plain language alongside the SQL or code that generates them • An interactive dashboard or report with drill-down capability and export options (PDF/CSV) • Deployment guide a...
...to work closely with stakeholders and convert business requirements into impactful reporting solutions. Key Responsibilities • Design, develop, and maintain interactive dashboards and reports using Power BI • Build scalable data models, DAX calculations, and Power Query transformations • Develop KPI dashboards, drill-through/drill-down reports, and executive MIS reports • Manage and optimize ETL workflows and data transformation processes • Perform report optimization for performance, usability, and scalability • Implement and manage Row-Level Security (RLS) and workspace governance • Work with SQL databases, Azure SQL, Synapse, APIs, and Excel integrations • Conduct business analysis, forecasting, trend analysis, and performance tr...
I need an experienced Power BI freelancer to step in and polish our curre...underlying data models—our models work, but they need tighter relationships, cleaner measures, and smarter DAX. • Maintain secure access by configuring Row-Level Security. • If the remodel uncovers gaps, design light ETL steps to reshape incoming tables before they hit the model. I will provide current .pbix files, schema diagrams, and a quick brief on the key business metrics we track. Success is a set of dashboards that refresh reliably, render quickly, and tell a clearer story to our stakeholders. Tools and skills you’ll lean on: Power BI Desktop/Service, DAX, SQL Server, ETL scripting, and best-practice data visualization. Remote delivery is fine; we can connect over ...
In this article, you learn how to create and configure a Zeppelin instance on an EC2, and about notebook storage on S3, and SSH access.