# cloudxlab.com > AI-optimized mirror of cloudxlab.com containing 24 pages totalling 310 words of clean markdown content, structured data, and semantic HTML. Original source: https://cloudxlab.com/. Last updated: 2026-06-15T08:10:59.335Z. Each page is available as HTML (with JSON-LD structured data) and Markdown (text-only, ideal for LLMs and RAG). ## Homepage - [CloudxLab Consulting | Enterprise AI & Data Engineering Solutions](/content/site-root.html): CloudxLab provides end-to-end consulting in Artificial Intelligence, Precision Data Engineering, and Distributed Systems for Fortune 500 companies globally. (54 words) ## Articles & Blog Posts - [ Explore Free Guided Projects on AI and Big Data | CloudxLab | CloudxLab ](/content/projects/explore/index.html): Explore Free Guided projects from top universities and industry leaders. Learn Machine Learning, Deep Learning and Big Data by doing hands-on. (1 words) - [ FAQ - I'm an instructor. How should I provide CloudxLab to my students? | CloudxLab ](/content/faq/12/im-an-instructor-how-should-i-provide-cloudxlab-to-my-students.html): Please sign up here as an instructor and we will provide you the details. (1 words) - [ Connect From Any Device | CloudxLab ](/content/features/connect-from-any-device/index.html): Practice Hadoop, Spark and Big Data technologies from any device. All you need is the internet access to CloudxLab. Practice right from your browser. (1 words) - [Embark on a transformative journey into the world of Artificial Intelligence and Machine Learning with CloudxLab's comprehensive courses. Dive deep into cutting-edge technologies and emerge as a proficient AI professional ready to tackle real-world challenges. Our expert-led programs offer hands-on experience and personalized mentorship, ensuring you master the skills demanded by today's top global firms. Join us and unlock your potential in the AI landscape, shaping the future with every line of code you write.](/content/artificial-intelligence-courses/index.html): Best Artificial Intelligence Courses: Empowering your Future in AI (102 words) - [Affiliate Program | CloudxLab](/content/affiliate/program/index.html): Promote CloudxLab on your blog, course page or website and earn commission on every sale. (1 words) - [ FAQ - Which hadoop distribution do you provide? | CloudxLab ](/content/faq/4/which-hadoop-distribution-do-you-provide/index.html): We provide the Hortonworks Data Platform. The Hadoop version on the cluster is 2.7. Please find the version of all the software components installed on CloudxLab here. (1 words) - [ FAQ - I have some more questions. Can I talk to someone? | CloudxLab ](/content/faq/13/i-have-some-more-questions-can-i-talk-to-someone/index.html): Absolutely! Please contact us here. You can also reach us anytime on our 24/7 support helpline by calling us on +918049202224 (1 words) - [ FAQ - What is your refund policy? | CloudxLab ](/content/faq/10/what-is-your-refund-policy/index.html): If you are unhappy with the product for any reason, let us know within 7 days of purchasing or upgrading your account, and we'll cancel your account and issue a full refund. Please contact us at reachus@cloudxlab.com to request a refund within the stipulated time. We will be sorry to see you go though! (1 words) - [ FAQ - Do I get a dedicated cluster of my own? | CloudxLab ](/content/faq/14/do-i-get-a-dedicated-cluster-of-my-own/index.html): CloudxLab is a shared cluster where you will be sharing resources with the other users. For dedicated cluster requests for multiple users reachus@cloudxlab.com. (1 words) - [Indian Institute of Technology Roorkee](/content/learn/index.html): Explore IIT Roorkee AI programs by CloudxLab. Choose between a 10-month PG Certificate in AI & Machine Learning or a 3-month Executive Certificate in Applied AI. IIT Roorkee certification, live classes, hands-on labs. (133 words) - [ FAQ - Will I get support? | CloudxLab ](/content/faq/8/will-i-get-support/index.html): Yes! Please feel free to ask your questions on CloudxLab forum and our community and team of experts will answer your questions. We believe forum will add better perspectives, ideas, and solutions to your questions. (1 words) - [ FAQ - What technologies can I practice on CloudxLab? | CloudxLab ](/content/faq/3/what-technologies-can-i-practice-on-cloudxlab/index.html): The tools and components available in the cluster include Hadoop, Spark, Kafka, Hive, Pig, HBase, Oozie, ZooKeeper, Flume, Sqoop, Mahout, R, Linux, Python, Scala, MongoDB, NumPy, SciPy, Pandas, Scikit-learn etc. Again, if you are looking for other tools please contact us at reachus@cloudxlab.com. (1 words) - [ FAQ - What are the limits on the usage of lab or what is the fair usage policy (FUP)? | CloudxLab ](/content/faq/6/what-are-the-limits-on-the-usage-of-lab-or-what-is-the-fair-usage-policy-fup.html): The CloudxLab cluster is used for educational and PoC purposes. The reason we are able to provide the cluster at a very low cost is that we are able to share the systems. The system resources are limited. If you try to use more resources, it is going to hurt other users. We have been avoiding putting hard limits on the resources consumption because we do not want to put roadblocks to the learning path to our users. Here are the limits as per the fair usage policy: HDFS - We provide 4.5 GB of storage space on HDFS with the replication factor of 3. That means if the replication factor is 3, you can store up to 1.5 GB data. And if the replication factor is 1, you can store up to 4.5 GB data. This is a hard limit meaning the HDFS will throw an error if you want to go beyond that storage. Local Storage on the Linux console - The allowed storage is 3 GB on the web console in your home directory. If you exceed this 3GB quota, you are given a 7 day grace period to reduce your usage to less than 3GB. During this grace period you can have a maximum of 4GB of data in your home directory. Once the grace period expires or you exceed the 4GB limit, whichever is earlier, you will no longer be able to create any new files in your home directory, and also your Jupyter server will stop working until you reduce your usage to less than 3GB. Also, our scripts keep observing the storage consumption. Our bots will automatically delete your data if your storage is more than 4GB. To clean the unnecessary files please follow the instructions given here. Hive - Please ensure that you do not create too many databases in Apache Hive. The permissible number of databases in the Apache Hive is one. RAM - Please ensure that your programs are not consuming memory (RAM) beyond 2GB. This hurts the other users. Our bots will automatically kill your processes if your RAM usages are more than 2GB. Duration - Please do not run a long process such as the Hive, pyspark, spark-shell, or Jupyter notebook. Your process will be killed by our bots if 1) It is running for more than 3 hours 2) Your notebook is idle for more than 60 mins, 3) You are using more than one YARN container at a time. While Hive, pyspark, spark-shell consume the containers from YARN, the Jupyter notebook consumes the local memory. CPU - Please do not run CPU-intensive tasks such as bitcoin mining or an infinite loop. Bandwidth - Please do not download more than 5 GB of data a month. MySQL - Please note that you will not be able to create new databases in MySQL. In MySQL, it becomes difficult to manage if we are allowing everyone to create databases. Please note that violating these terms is an offense and your account might get disabled in case of an offense. (1 words) - [ FAQ - How many nodes in the cluster? | CloudxLab ](/content/faq/5/how-many-nodes-in-the-cluster/index.html): Currently, we have 5 nodes in the cluster. We automatically scale up and down based on the cluster load. Three nodes have 8 cores and 32 GB RAM and the other two nodes have 16 cores and 60 GB RAM each depending on the services running on them. (1 words) - [ FAQ - Are there sample datasets in the cluster? | CloudxLab ](/content/faq/7/are-there-sample-datasets-in-the-cluster/index.html): We have tried to keep everything available so you can start practicing without delay. So, yes, we do provide sample datasets. Please find the list of available datasets here (1 words) - [ Student Login | CloudxLab ](/content/refer-friend/index.html) (1 words) - [ Schedule A Demo | CloudxLab ](/content/schedule-demo/index.html): Please leave your information in the form below, and we will contact you to set up an online demo of CloudxLab. (1 words) - [ Centralized Data sets | CloudxLab ](/content/features/centralized-datasets/index.html): Practice Hadoop, Spark and Big Data technologies in real world data sets already provided in HDFS. No more rummaging around for data to run a test on. (1 words) - [ No Installation and Compatibility Issues | CloudxLab ](/content/features/no-installation/index.html): Practice Hadoop, Spark and Big Data technologies without worrying about installation and compatibility issues. Practice right from your browser without installing any software. (1 words) - [ Connect From Anywhere | CloudxLab ](/content/features/connect-from-anywhere/index.html): Practice Hadoop, Spark and Big Data technologies anytime, anywhere in the world. Practice right from your browser. (1 words) - [ Learn Through Practice | CloudxLab ](/content/features/learn-through-practice/index.html): Practice makes perfect. Learn and practice Hadoop, Spark and Big Data technologies in real time cluster anywhere, anytime. (1 words) - [ Real World Experience | CloudxLab ](/content/features/real-world-experience/index.html): Practice Hadoop, Spark and Big Data technologies in real world cluster and get hands-on experience in real production environment. (1 words) - [ Contact Us | CloudxLab ](/content/reach-us-queries/index.html): Have question regarding anything on CloudxLab or would just like to have a chat with us? Reach us on phone or email. (1 words) ## Resources - [Full Page Index](/index.html): Browse all cached pages with rich metadata - [About This Cache](/content/about.html): Methodology, technical details, and usage guidelines - [XML Sitemap](/sitemap.xml): Machine-readable sitemap for crawler discovery - [Robots.txt](/robots.txt): Crawler directives