Are you a Windows/.NET developer and willing to learn big data concepts and tools in your Windows?

If yes, you can follow the links below to install them in your PC. The installations are usually easier to do in Linux/UNIX but they are not difficult to implement in Windows either since they are based on Java.

Installation guides

All the following documents are based on Windows 10. The steps should be the same in other Windows environments though some of the screenshots may be different.

Install Zeppelin 0.7.3 in Windows

Install Hadoop 3.0.0 in Windows (Single Node)

Install Spark 2.2.1 in Windows

Install Apache Sqoop in Windows

Configure Hadoop 3.1.0 in a Multi Node Cluster

Apache Hive 3.0.0 Installation on Windows 10 Step by Step Guide

Learning tutorials

Use Hadoop File System Task in SSIS to Write File into HDFS
Invoke Hadoop WebHDFS APIs in .NET Core

Write and Read Parquet Files in Spark/Scala

Write and Read Parquet Files in HDFS through Spark/Scala

Convert String to Date in Spark (Scala)

Read Text File from Hadoop in Zeppelin through Spark Context

Connecting Apache Zeppelin to your SQL Server

Load Data into HDFS from SQL Server via Sqoop

Default Ports Used by Hadoop Services (HDFS, MapReduce, YARN)

Connect to SQL Server in Spark (PySpark)

Implement SCD Type 2 Full Merge via Spark Data Frames

Password Security Solution for Sqoop

PySpark: Convert JSON String Column to Array of Object (StructType) in Data Frame

Spark - Save DataFrame to Hive Table

Copy Files from Hadoop HDFS to Local

Data Partitioning in Spark (PySpark) In-depth Walkthrough

Data Partitioning Functions in Spark (PySpark) Deep Dive

Read Data from Hive in Spark 1.x and 2.x

Get the Current Spark Context Settings/Configurations

PySpark - Fix PermissionError: [WinError 5] Access is denied

Configure a SQL Server Database as Remote Hive Metastore

Connect to Hive via HiveServer2 JDBC Driver

I will be constantly updating my blog with tutorials. Feel free to subscribe this blog (RSS).

info Last modified by Raymond at 2 years ago * This page is subject to Site terms.

More from Kontext

local_offer spark local_offer pyspark

visibility 3880
thumb_up 0
access_time 12 months ago

When creating Spark date frame using schemas, you may encounter errors about “field **: **Type can not accept object ** in type <class '*'>”. The actual error can vary, for instances, the following are some examples: field xxx: BooleanType can not accept object 100 in type ...

open_in_new Spark + PySpark

local_offer hive

visibility 1645
thumb_up 0
access_time 2 years ago

Since Hive 3.x, new authentication feature for HiveServer2 client is added. When starting HiveServer2 service (Hive version 3.0.0), you may encounter errors like: ‘HiveServer2 metastore.RetryingMetaStoreClient: RetryingMetaStoreClient trying reconnect as [username]  (auth:S...

open_in_new Hadoop

local_offer linux local_offer WSL local_offer ubuntu

visibility 5569
thumb_up 3
access_time 2 years ago

This page shows how to install Windows Subsystem for Linux (WSL) system on a non-system drive manually. Enable Windows Subsystem for Linux system feature Open PowerShell as Administrator and run the following command to enable WSL feature: Enable-WindowsOptionalFea...

open_in_new Tools

local_offer kontext

visibility 69
thumb_up 0
access_time 2 years ago

In the past months, this website has been using the following Email address to delivery all the notification messages to the website users such as registration confirmation email, comment email and so on. no-reply[at] However, I recently fou...

open_in_new Kontext Information

info About author

comment Comments (2)

comment Add comment

Please log in or register to comment.

account_circle Log in person_add Register

Log in with external accounts


Yes, you can load your text file into hdfs via CLI, WebHDFS api or any other tools/programming that supports this. You can then do transformations using tools like Apache Beam, Spark or notebooks (Zeppeline or Jupyter), etc.These tools also can then write into sql server database through ODBC/JDBC or native SQL Server drivers.


person Rajesh access_time 2 years ago
Re: Install Big Data Tools (Spark, Zeppelin, Hadoop) in Windows for Learning and Practice

Can we load date from text file to HDFS and do calculations/corrections and load the data to SQL SERVER DB table

reply Reply
account_circle Rajesh

Can we load date from text file to HDFS and do calculations/corrections and load the data to SQL SERVER DB table

reply Reply

Dark theme mode

Dark theme mode is available on Kontext.

Learn more arrow_forward

Kontext Column

Created for everyone to publish data, programming and cloud related articles. Follow three steps to create your columns.

Learn more arrow_forward