<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: Benard Otieno</title>
    <description>The latest articles on DEV Community by Benard Otieno (@benard_otieno_254).</description>
    <link>https://dev.to/benard_otieno_254</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F3841609%2F9e496ffd-ee20-4209-89ab-8f6fe772bf29.jpg</url>
      <title>DEV Community: Benard Otieno</title>
      <link>https://dev.to/benard_otieno_254</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/benard_otieno_254"/>
    <language>en</language>
    <item>
      <title>Parametric vs Non-parametric tests</title>
      <dc:creator>Benard Otieno</dc:creator>
      <pubDate>Sun, 13 Sep 2026 18:55:40 +0000</pubDate>
      <link>https://dev.to/benard_otieno_254/parametric-vs-non-parametric-tests-3na</link>
      <guid>https://dev.to/benard_otieno_254/parametric-vs-non-parametric-tests-3na</guid>
      <description>&lt;p&gt;&lt;strong&gt;Introduction&lt;/strong&gt;&lt;br&gt;
Statistical test is one of the first road maps when decieding whether to use parametric or non-parametric tests. The choice affects how valid your results are, that is the power of your test and the confidence you place in your conclusions. Choosing whether to use a parametric or a no-parametric test has a great impact on your prediction either negatively or positively.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fcg4x3sk1xdc1dv2zf10k.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fcg4x3sk1xdc1dv2zf10k.png" alt=" " width="628" height="426"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Parametric Test&lt;/strong&gt;&lt;br&gt;
Parametric test assumes that a polpulation data follows a specific probability distribution, commonly known as normal distribution. This distribution can be described by a set of parameters namely; mean and variance. Since the distribution of parametric tests is assumed to be known, the test tends to be be statistically powerful and can extract more information from the data.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Core Assumptions of Parametric Tests&lt;/strong&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Normality:&lt;/strong&gt; The data is normally distributed or approximately normally distributed.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Homogeneity of variance:&lt;/strong&gt; Groups that are bieng compared have similar or approximately similar variance.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Interval or ratio scale data:&lt;/strong&gt; Meanignful spacing of data.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Independence of observations:&lt;/strong&gt; One data point does not have an influence on the others.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;strong&gt;Types of Parametric tests&lt;/strong&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;&lt;p&gt;&lt;strong&gt;One sample t-test&lt;/strong&gt;&lt;br&gt;
Used to compare one sample mean to a known value. For instance, the standard weight of bred is 400grams, if we have 10 loaves of bread, we add their weights and get the mean of the 10 loaves of bread and compare it with the standard 400grams.&lt;/p&gt;&lt;/li&gt;
&lt;li&gt;&lt;p&gt;&lt;strong&gt;Independant t-test&lt;/strong&gt;&lt;br&gt;
Compares the mean of two independant groups. The independant variable is always a categorical variable while the dependant variable is a continous variable. For instance a drug test on two sets of five people, the first five people are given 300mg of panadol while the second group is given 500mg of panadol at the same time to see the of both groups after the test.&lt;/p&gt;&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;3.&lt;strong&gt;Paired t-test&lt;/strong&gt;&lt;br&gt;
Compares the mean of two related samples. For instance a drug test on five people, for the first they are given 300mg of panadol while for the second  test they are given 500mg of panadol at a different time to see the results before and after. The independant variable is always a categorical variable while the dependant variable is a continous variable.&lt;/p&gt;

&lt;p&gt;4.&lt;strong&gt;One-way ANOVA&lt;/strong&gt;&lt;br&gt;
Compare means of more than two groups, Compare average satisfaction scores across three product tiers. Similarly, the independant variable is always a categorical variable while the dependant variable is a continous variable.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fwomp0r95xli1fghri4yt.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fwomp0r95xli1fghri4yt.png" alt=" " width="800" height="406"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;5.&lt;strong&gt;Pearson correlation&lt;/strong&gt;&lt;br&gt;
Measure linear relationship between two continuous variables. Used for numerical variables.&lt;/p&gt;

&lt;p&gt;6.&lt;strong&gt;Linear regression&lt;/strong&gt;&lt;br&gt;
Model a continuous outcome using predictors&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Non-parametric test&lt;/strong&gt;&lt;br&gt;
Non-parametric test has little to no assumption about the distribution. Many non-parametric tests convert data into ranks instead of directly working with raw data. This makes them resistant to skewed distribution, non-normal distributions and outliers. Non-parametric tests are suitable when:&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;You have an ordinal data instead of continous.&lt;/li&gt;
&lt;li&gt;Your data has extreme outliers or heavily skewed.&lt;/li&gt;
&lt;li&gt;You have a categorical data and you're testing for association instead of means.&lt;/li&gt;
&lt;li&gt;The sample size of your data is small, making it difficult to validate normality.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;&lt;strong&gt;Types Non-Parametric Tests&lt;/strong&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Mann-Whitney U test&lt;/strong&gt;
This is the alternative for Independant t-test that does not qualify the parametric tests assumptions.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fxgffrdyiuis45o7i0ra4.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fxgffrdyiuis45o7i0ra4.png" alt=" " width="799" height="468"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Wilcoxon signed-rank test&lt;/strong&gt;
This is the alternative for paired t-test that does not meet the parametric tests assumptions.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fpawjrppzx5xy4n0qnhlb.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fpawjrppzx5xy4n0qnhlb.png" alt=" " width="799" height="511"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Kruskal-Wallis test&lt;/strong&gt;
This is the alternative for One-way ANOVA that does not meet the parametric tests assumptions.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Ffthx49jrm6ri83jp5d8l.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Ffthx49jrm6ri83jp5d8l.png" alt=" " width="800" height="356"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Spearman's rank correlation&lt;/strong&gt;
This is the alternative for Pearson correlation that does not meet the parametric tests assumptions.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Foxnxi5ugr7c2666cfcoz.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Foxnxi5ugr7c2666cfcoz.png" alt=" " width="800" height="682"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Chi-square test of independence&lt;/strong&gt;
This is a non-parametric test used to check for relationship between categorical variables.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F9hktqrq4rsootwyzmz9m.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F9hktqrq4rsootwyzmz9m.png" alt=" " width="800" height="443"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;The Table below shows side by side comparison of parametric and non-parametric tests.&lt;/strong&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F9gduz6b4ej2s8q79phw2.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F9gduz6b4ej2s8q79phw2.png" alt=" " width="800" height="297"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Conclusion&lt;/strong&gt;&lt;br&gt;
Choosing between parametric and non-parametric tests does not really matter whether you chose the correct on or not, It solely relies on the nature of data. When assumption are held, parametric tests offer more power in terms of statistical analysis. However, they can produce results that are misleading when applied to data that is ordinal, small and has a skewed distribution. Therefore non-parametric tests offer alternatives when data does not meet the assumptions. In conclusion, before running any tests, ensure to check the assumptions in-order to produce correct results.&lt;/p&gt;

</description>
      <category>data</category>
      <category>datascience</category>
      <category>science</category>
    </item>
    <item>
      <title>Data Distribution and their impact in Data Science.</title>
      <dc:creator>Benard Otieno</dc:creator>
      <pubDate>Sun, 13 Sep 2026 08:41:45 +0000</pubDate>
      <link>https://dev.to/benard_otieno_254/data-distribution-and-their-impact-in-data-science-2e60</link>
      <guid>https://dev.to/benard_otieno_254/data-distribution-and-their-impact-in-data-science-2e60</guid>
      <description>&lt;p&gt;&lt;strong&gt;Introduction&lt;/strong&gt;&lt;br&gt;
Datasets show distribution and it's shape. Distribution is the pattern that explains hows values are spread from a variable. Data distribution is important in data science since it influences most of the decisions that follows. For instance, the valid statistical test to use, the machine learning algorithm that will perform well, hoe to handle extreme outliers and how to honestly communicate results.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;What is Data Distribution&lt;/strong&gt;&lt;br&gt;
Data distribution is a graphical represenattion that explains the frequency of occurennces of each valuein a dataset. It is mostlyb visualised using tools like boxplots, histograms and density plots.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Why Distributions Matter in Data Science&lt;/strong&gt;&lt;/p&gt;

&lt;p&gt;**&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;&lt;p&gt;Choosing the right statistical tests&lt;br&gt;
**&lt;br&gt;
Many hypothesis tests assume normal distribution of data ansd therefore applying heavily skewed datawithout checking the assumptions leads to false conclusion. Alternatively, Non-parametric tests exists for the cases where there is no normal distribution.&lt;br&gt;
**&lt;/p&gt;&lt;/li&gt;
&lt;li&gt;&lt;p&gt;Outliers Detection&lt;br&gt;
**&lt;br&gt;
Outliers are data points that have statistical anomaly, in simple terms outliers are data points that significantly deviates from other data points in the dataset. They can be brought about by natural variation or measurement errors and they can have a negative impact on the statistical analysis and machine learning model.&lt;/p&gt;&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fntj6t2prubrv6zqfwvds.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fntj6t2prubrv6zqfwvds.png" alt=" " width="800" height="82"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;**&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;Model Selection
**
Some models are built around certain distributional assumptions:&lt;/li&gt;
&lt;/ol&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Linear Regresssion;&lt;/strong&gt; This model assumes normal distribution.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Poisson Regression;&lt;/strong&gt; Designed to count data that follows a poisson distribution.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Naive Bayes classifiers;&lt;/strong&gt; Mostly assumes data following Gaussian distribution&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;&lt;strong&gt;## Common types of Data Distributions&lt;/strong&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;1. Normal Distribution&lt;/strong&gt;&lt;br&gt;
This is the most common type of distribution also known as Bell-Curve. It is symmetric around the mean where most of the values cluster near the center.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Ff0nabhh6zxlmac4g392k.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Ff0nabhh6zxlmac4g392k.png" alt=" " width="800" height="628"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;2. Uniform Distribution&lt;/strong&gt;&lt;br&gt;
In uniform distribution every outcome within a range is distributed equally.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fe43gxy826qkmjsppgffe.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fe43gxy826qkmjsppgffe.png" alt=" " width="575" height="378"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;3. Right Skewed Distribution&lt;/strong&gt;&lt;br&gt;
In Right skewed distribution, the tail stretches to the right, in most ocassion the mean is higher than the median. This is brought about by outliers that dragging the mean to the right. It is also known as a positive distribution.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fa0ogehm825ew4s7gbyhe.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fa0ogehm825ew4s7gbyhe.png" alt=" " width="799" height="526"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;4. Left Skewed Distribution&lt;/strong&gt;&lt;br&gt;
In Left skewed distribution, the tail stretches to the left, in most ocassion the median is higher than the mean. This is brought about by outliers that dragging the mean to the left. It is also known as a negative distribution.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F1c2t548edqhent5rdnpi.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F1c2t548edqhent5rdnpi.png" alt=" " width="800" height="519"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;5. Binomial Distribution&lt;/strong&gt;&lt;br&gt;
Explains the number of successes in a fixed number of yes or no trials. Mostly used in quality control and A/B testing conversion rate analysis. &lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fac6rdso81qf7lblwu2qa.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fac6rdso81qf7lblwu2qa.png" alt=" " width="692" height="403"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Practical Tips for Working with Distributions&lt;/strong&gt;&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;Visualize your data before modelling.&lt;/li&gt;
&lt;li&gt;Test for normality&lt;/li&gt;
&lt;li&gt;Do not fix normality where it does not belong&lt;/li&gt;
&lt;li&gt;Check for multimodal distribution&lt;/li&gt;
&lt;li&gt;After pre-processing, re-check distribution&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;&lt;strong&gt;Visualizing Distributions&lt;/strong&gt;&lt;br&gt;
Visualization is one of the easiest ways to check how your data is distributed. Some of the ways we can visualise to check for distribution include:&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Histogram&lt;/strong&gt;
Visualizes countinous data  into intervals known as bins and in each bin, frequency of occurence is counted. Hence showing the overal distribution of data.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F9foss7xe0c8c0vpd4j1u.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F9foss7xe0c8c0vpd4j1u.png" alt=" " width="800" height="800"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Box Plot&lt;/strong&gt;
Visualizes distribution of numeric variables using five key summary statistics. &lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F2vl2vnr1jr7f9p3g8pg9.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2F2vl2vnr1jr7f9p3g8pg9.png" alt=" " width="799" height="383"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fcpid69q1vima253k5n3r.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Farticles%2Fcpid69q1vima253k5n3r.png" alt=" " width="800" height="373"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;The box spans the middle 50% of the data, from Q1 which is 25th percentile to Q3 which is 75th percentile. &lt;/li&gt;
&lt;li&gt;The 50th percentile is marked by the line inside the box which shows the median and not mean, hence making boxplot resistant to skew and outliers&lt;/li&gt;
&lt;li&gt;The whiskers extend to the smallest and largest values that are still within 1.5 × IQR of the box edges.&lt;/li&gt;
&lt;li&gt;Points that are beyond the whiskers are outliers &lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;&lt;strong&gt;Conlusion&lt;/strong&gt;&lt;br&gt;
Data distributions is a practical lens that shapes nearly every step of the data science workflow, from exploratory data analysis(EDA) to feature engineering to model selection and evaluation. A data scientist who understands the distribution of their data will make better decisions about which tools to use, avoid misleading conclusions, and build models that generalize more reliably to the real world. Before running a single test or training a single model, the first and most valuable question to ask is often the simplest one: &lt;strong&gt;_what does this data actually look like _&lt;/strong&gt;&lt;/p&gt;

</description>
    </item>
    <item>
      <title>How to publish a Power BI report and embed it in a website.</title>
      <dc:creator>Benard Otieno</dc:creator>
      <pubDate>Sat, 04 Apr 2026 12:07:08 +0000</pubDate>
      <link>https://dev.to/benard_otieno_254/how-to-publish-a-power-bi-report-and-embed-it-in-a-website-2n77</link>
      <guid>https://dev.to/benard_otieno_254/how-to-publish-a-power-bi-report-and-embed-it-in-a-website-2n77</guid>
      <description>&lt;p&gt;&lt;strong&gt;Introduction to Power BI&lt;/strong&gt;&lt;br&gt;
Power BI is a data analytics tool used to turn raw data into interactive reports and dashboard. It enables users to obtain insights, visualize data and make decisions that are data driven.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Publishing and Embeding Process&lt;/strong&gt;&lt;br&gt;
In Power BI, this process involves workspace creation in a Power BI service, thereafter you upload and publish the Power BI report, before you generate and embed code and lastly embedding the report to the website.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;1. Creating a Workspace&lt;/strong&gt;&lt;br&gt;
In Power BI, a worksapce is a collaborative shared enviroment for hosting dashboards, reports and managing datasets. Below are steps and examples of how to create a workspace.&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;&lt;p&gt;Step 1: Open your browser and click this website &lt;a href="https://app.powerbi.com/home?experience=power-bi" rel="noopener noreferrer"&gt;https://app.powerbi.com/home?experience=power-bi&lt;/a&gt;&lt;/p&gt;&lt;/li&gt;
&lt;li&gt;&lt;p&gt;Step 2: Sign in to your account.&lt;br&gt;
&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F582l5qf2ud9c5qs5kuox.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F582l5qf2ud9c5qs5kuox.png" alt=" " width="605" height="462"&gt;&lt;/a&gt;&lt;/p&gt;&lt;/li&gt;
&lt;li&gt;&lt;p&gt;Step 3: Once you're logged in, go to the left navigation paneand click workspaces.&lt;/p&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fjc42wcptu0gunljeuytw.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fjc42wcptu0gunljeuytw.png" alt=" " width="800" height="321"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;&lt;p&gt;Step 4: After opening the workspaces, click +New Workspace&lt;br&gt;
&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F02sfc0o4eoqtdlo6nx57.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F02sfc0o4eoqtdlo6nx57.png" alt=" " width="401" height="143"&gt;&lt;/a&gt;&lt;/p&gt;&lt;/li&gt;
&lt;li&gt;&lt;p&gt;Step 5: Lastly enter the workspace name and description to the pop up then click apply to save.&lt;/p&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;&lt;strong&gt;2. Uploading and Publishing a Power BI report&lt;/strong&gt;&lt;br&gt;
Open the Power BI report in your computer then click Publish which is located at the top right.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fba2jh5nfj9tf64o9fv7g.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fba2jh5nfj9tf64o9fv7g.png" alt=" " width="89" height="135"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Thereafter select your workspace and the report will be uploaded and the and becomes accessible online.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F8oe7yfzvbfggyia0hyni.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F8oe7yfzvbfggyia0hyni.png" alt=" " width="755" height="579"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fk2okjdat6e0lii68oxys.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fk2okjdat6e0lii68oxys.png" alt=" " width="659" height="398"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;3. Generating the embed code&lt;/strong&gt;&lt;br&gt;
Power BI enables us to embed reports via iframe code. To begin you need to open the report in the Power BI service, then click file, hover down to embed report and lastly click website or portal to create the embed code.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fk5tnhyhb10p5fk0lixtu.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fk5tnhyhb10p5fk0lixtu.png" alt=" " width="800" height="365"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fgbxyiyqf6outekigmy2x.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fgbxyiyqf6outekigmy2x.png" alt=" " width="596" height="495"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fyto9seigcslgcydqrnzv.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fyto9seigcslgcydqrnzv.png" alt=" " width="800" height="424"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;4. Embedding the Report on a Website&lt;/strong&gt;&lt;br&gt;
Since we now have the code, we can easily intergrate it in the HTML code. Open this website in your browser &lt;a href="https://www.w3schools.com/html/" rel="noopener noreferrer"&gt;https://www.w3schools.com/html/&lt;/a&gt; and copy the html code.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Ff5t0s5xqx8q6vmw968pn.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Ff5t0s5xqx8q6vmw968pn.png" alt=" " width="456" height="380"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Open a text editor and go back to the desktop a create a folder with a name of your choice. Go back to the text editor and and open a new folder and select the folder you created.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fthikl4emzgy1i6hsh02k.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fthikl4emzgy1i6hsh02k.png" alt=" " width="508" height="707"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fs10bkqom5cho9jwim0wp.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fs10bkqom5cho9jwim0wp.png" alt=" " width="800" height="528"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Now in the text editor go to the folder you created and create a new file with a name of your choice, then paste the HTML code that was earlier copied. Finally, in the text editor delete line nine of the html code and replace it with the embed code and save then open the folder you created in the desktop and open the file in the folder which leads to this page below.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fcwsvzjenvaapp530h7tb.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fcwsvzjenvaapp530h7tb.png" alt=" " width="800" height="369"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Lastly sign in to the page and your report will be visible.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fij0ivl5ryx1ux71ht7ed.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fij0ivl5ryx1ux71ht7ed.png" alt=" " width="800" height="366"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Conclusion&lt;/strong&gt;&lt;br&gt;
Publishing a power BI report and embedding it enables the sharing of insights not only on the Power BI platform. If there is proper optimization and security , embedded Power BI reports has the potential to become a powerful feature of modern data-driven websites.&lt;/p&gt;

</description>
      <category>analytics</category>
      <category>microsoft</category>
      <category>tutorial</category>
      <category>webdev</category>
    </item>
    <item>
      <title>Understanding Data Modeling in Power BI.</title>
      <dc:creator>Benard Otieno</dc:creator>
      <pubDate>Mon, 30 Mar 2026 14:36:50 +0000</pubDate>
      <link>https://dev.to/benard_otieno_254/understanding-data-modeling-in-power-bi-48ae</link>
      <guid>https://dev.to/benard_otieno_254/understanding-data-modeling-in-power-bi-48ae</guid>
      <description>&lt;p&gt;&lt;strong&gt;Data Modelling&lt;/strong&gt;&lt;br&gt;
Data modelling involves creating a visual diagram of data to show how it can be organized, structered and analysed within a database. In Power BI, data modelling enables ease in report creation, accuracy in calculations and efficient performance.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;SQL JOINS&lt;/strong&gt;&lt;br&gt;
SQL JOINS is used to combine data in multiple tables based on the related columns between them. They are important in relational database since it enables us to retrieve data with unified view that is mostly spread across multiple tables.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;INNER JOIN&lt;/strong&gt;&lt;br&gt;
INNER JOIN is an SQL operation that enables combining rows from multiple tables with respect to a related column.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;LEFT JOIN&lt;/strong&gt;&lt;br&gt;
LEFT JOIN retains rows from the left side of the table, and combines rows that are matching from the right side of the table.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;RIGHT JOIN&lt;/strong&gt;&lt;br&gt;
RIGHT JOIN merges two tables by retaining data from the right ide of the table and only matching data from the left side of the table.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;FULL OUTER JOIN&lt;/strong&gt;&lt;br&gt;
FULL OUTER JOIN merges two tables by retaining all the data from both tables, regardless of whether they are matching.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;LEFT ANTI JOIN&lt;/strong&gt;&lt;br&gt;
LEFT ANTI JOIN returns rows from the left side of the table that do not have a match on the right side of the table.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;RIGHT ANTI JOIN&lt;/strong&gt;&lt;br&gt;
RIGHT ANTI JOIN returns rows from the right side of the table that do not have a match on the left side of the table.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;How to access JOINS in a Power BI query&lt;/strong&gt;&lt;br&gt;
To access JOINS, go to HOME then click transform data &amp;gt; open power query editor &amp;gt; choose second table &amp;gt; select matching columns, then select the join type and lastly click ok.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Relationships in Power BI&lt;/strong&gt;&lt;br&gt;
Relationships connects tables, without merging them physically, enabling records from multiple places to be used jointly in reports. There are four types of relationships namely;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;One-to-many&lt;/strong&gt;&lt;br&gt;
One-to-many connects a lookup table that contains values that are unique , to a table, where the values repeat.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Many-to-many&lt;/strong&gt;&lt;br&gt;
Many-to-many normally occur in a relationship when both tables contain duplicate values , enabling multiple data in one table to match multiple data in another.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;One-to-one&lt;/strong&gt;&lt;br&gt;
One-to-one connects two different tables whereby two unique values from the different tables are matched.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Cardinality&lt;/strong&gt;&lt;br&gt;
Cardinality explains the kind of the relationship between table A and table B, describing how unique values in one column relate to values in another column.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Cross-filter Direction&lt;/strong&gt;&lt;br&gt;
Cross-filter Direction determines the flow of filters in related tables, with options for single direction or both direction.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Active vs Inactive Relationships&lt;/strong&gt;&lt;br&gt;
Active relationships are used in visuals by default while Inactive relationships has to be activated using DAX.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Creating Relationships in Power BI&lt;/strong&gt;&lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;&lt;p&gt;Go to Model View&lt;/p&gt;&lt;/li&gt;
&lt;li&gt;&lt;p&gt;Drag a column from one table to the next&lt;/p&gt;&lt;/li&gt;
&lt;li&gt;&lt;p&gt;Lastly set to Cardinality or Cross-filter direction&lt;/p&gt;&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;strong&gt;Joins vs Relationships&lt;/strong&gt;&lt;br&gt;
In Joins, data is combined in one table, hence increasing the size of dataset, while in relationship, the performance is fast and efficient due to the schemas.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Fact and Dimension Tables&lt;/strong&gt;&lt;br&gt;
Fact tables has data that is quantitative and forms the star schema, while dimension tables hold descriptive data.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Data Modeling Schemas&lt;/strong&gt;&lt;br&gt;
A Schema is a framework and organization of data within in a model, which is significant for efficient data analysis.There are two types of schemas namely Star schema and Snowflake schema.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Star schema&lt;/strong&gt;&lt;br&gt;
Star schema organizes data into a central fact table and is surrounded by dimension tables.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Snowflake schema&lt;/strong&gt;&lt;br&gt;
Snowflake schema is an approach where dimension tables are normalized and broken into numerous tables that are related.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Flat Table&lt;/strong&gt;&lt;br&gt;
Flat Table is a table that contains all data, including descriptive and transactional facts.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Conclusion&lt;/strong&gt;&lt;br&gt;
Power BI seemed something complicated to me until I engaged in hands on learning that has made it easier for me to understand. Data medelling is a very important to learn since it is fundamental in analysing and reporting. &lt;/p&gt;

</description>
      <category>beginners</category>
      <category>data</category>
      <category>analytics</category>
      <category>datascience</category>
    </item>
    <item>
      <title>How Excel is Used in Real-World Data Analysis</title>
      <dc:creator>Benard Otieno</dc:creator>
      <pubDate>Thu, 26 Mar 2026 20:25:17 +0000</pubDate>
      <link>https://dev.to/benard_otieno_254/how-excel-is-used-in-real-world-data-analysis-pm7</link>
      <guid>https://dev.to/benard_otieno_254/how-excel-is-used-in-real-world-data-analysis-pm7</guid>
      <description>&lt;p&gt;Microsoft Excel is a tool used to collect, organize, analyse and visualize data. Various industries use excel to analyse data, for instance the businesses and marketing sector. Businesses use excel to keep track of their income, profit and losses, so as to see where they can improve on. On the other hand, sales and marketing sector use excel to analyse trend in sales and consumer data.&lt;br&gt;
Some of the excel features include example: &lt;/p&gt;

&lt;ol&gt;
&lt;li&gt;
*&lt;em&gt;Formulas *&lt;/em&gt;
&lt;/li&gt;
&lt;li&gt;AVERAGE, Which is used to calculate the average e.g. =AVERAGE(B2:B30)&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fl03l516gx9g199pixp7z.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fl03l516gx9g199pixp7z.png" alt=" " width="249" height="28"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;ul&gt;
&lt;li&gt;SUM, Which is used to addition of a range of numbers e.g SUM(B2:B30)&lt;/li&gt;
&lt;/ul&gt;

&lt;ol&gt;
&lt;li&gt;
Another feature used in excel also includes pivot table which is used to analyse big data. Example of a Pivot table&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fhnuqirgxgot4utupmdda.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fhnuqirgxgot4utupmdda.png" alt=" " width="514" height="204"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Pivort Table&lt;/strong&gt;&lt;br&gt;
Pivot tables is an excel tool used to analyse and simplify data without using formulas. To access the pivot table, click anywhere in the data and go to insert, then at the top right corner click pivot table then select New worksheet.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F2w6ybf6m1n0vutgw0tjy.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F2w6ybf6m1n0vutgw0tjy.png" alt=" " width="310" height="165"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Slicers&lt;/strong&gt;&lt;br&gt;
Slicers are used in excel to filter pivort tables and graphs. It is a easy to use since it helps display what is currently filtered by just clicking the button menus.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fvm32berkt109ivk3dakm.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Fvm32berkt109ivk3dakm.png" alt=" " width="268" height="338"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Graphs&lt;/strong&gt;&lt;br&gt;
Grphs are used in excel to show a visual representation of numerical data and analyse the data in more easier way to understand and intepret. Graphs enable the identifivation of trends, highlighting anomalies and comparing categories. Exaples of graphs include, line grapg and bar graph etc.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F1posrmssjd894972qqak.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F1posrmssjd894972qqak.png" alt=" " width="649" height="400"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F7o8b29rvsl11rbqhk24p.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2F7o8b29rvsl11rbqhk24p.png" alt=" " width="648" height="428"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Dashboard&lt;/strong&gt;&lt;br&gt;
A dashboard is a visual report that helps to simplify Key Perfomance Indicater through tables, charts and graphs that are interactive. Dashboard enables raw data to be transformed into meaningful insights, to help in decision making by enabling users to filter data through slicers for automatic update of data.&lt;/p&gt;

&lt;p&gt;&lt;a href="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Frjkpfcfnz4g4omw79lfj.png" class="article-body-image-wrapper"&gt;&lt;img src="https://media2.dev.to/dynamic/image/width=800%2Cheight=%2Cfit=scale-down%2Cgravity=auto%2Cformat=auto/https%3A%2F%2Fdev-to-uploads.s3.amazonaws.com%2Fuploads%2Farticles%2Frjkpfcfnz4g4omw79lfj.png" alt=" " width="800" height="298"&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;In Summary&lt;/strong&gt;&lt;br&gt;
Microsoft Excel is a significant tool for data analysis. Excel has the ability to do simple calculations, it also enables users to turn raw data into meaningful insights. Excel features such as Pivot Tables, and visualization tools like the graphs, individuals can significantly enhance their ability to analyze and interpret data.&lt;/p&gt;

</description>
      <category>analytics</category>
      <category>beginners</category>
      <category>data</category>
      <category>datascience</category>
    </item>
    <item>
      <title>Excel for beginners</title>
      <dc:creator>Benard Otieno</dc:creator>
      <pubDate>Thu, 26 Mar 2026 12:22:24 +0000</pubDate>
      <link>https://dev.to/benard_otieno_254/excel-for-beginners-2kmo</link>
      <guid>https://dev.to/benard_otieno_254/excel-for-beginners-2kmo</guid>
      <description></description>
      <category>beginners</category>
      <category>analyst</category>
    </item>
  </channel>
</rss>
