\documentclass[11pt]{article} %Gummi|065|=) \title{\textbf{Draft: Notes on the Pareto distribution}} \author{John Creighton\\ Nobody else} \date{} \usepackage{amsmath} \usepackage{hyperref} \usepackage[pdftex]{graphicx} \usepackage[utf8]{inputenc} \begin{document} \maketitle \section{Introduction} The Pareto distribution is a well knon model for wealth and income distribution and sometimes (e.g. Jordan Peterson) used as a model for varations in productivity.However, for this to be an acurate model of productivity in the tail of the distribution (AKA survival function) One must assume that one's pay is proportional to their productivity. This is a highly contentious claim and for instance would contradict Marx's labour theory of value. Under Marx's theory the capitalist as relativly productive as indicated by wealth and income but instead apears so due to taking a disproporationate amount of the surplus value that is produced by the collective enterprise. \section{Cross Dicipline Interest} The most important cross discipline relevance of the Pareto distribution is the tail behaviour. Near the tail the shape of the curve is obscured by an integrating constant which approaches one at infitinity. This asymptotic constant dispears in the survival function \cite{wikipediaSurvivalFunction}. \begin{equation} S(t)=P(\{T>t\})=\int _{t}^{\infty }f(u)\,du=1-F(t). \end{equation} which is more suitable for curve fitting both in a numerical or statsitcial sense if one is interested primarly in the tail behavior. The survival function is often written with a bar on top of the cumulative distribution function. This bar denotes logical negation. Formally the Pereto distrubtion is usually defined in terms of the survival function: \begin{equation} {\displaystyle {\overline {F}}(x)=\Pr(X>x)={\begin{cases}\left({\frac {x_{\mathrm {m} }}{x}}\right)^{\alpha }&x\geq x_{\mathrm {m} },\\1&x2$. This can be scene from it's probability density function: \begin{equation} f_X(x)= \begin{cases} \frac{\alpha x_\mathrm{m}^\alpha}{x^{\alpha+1}} & x \ge x_\mathrm{m}, \\ 0 & x < x_\mathrm{m}. \end{cases} \end{equation} which is a power law distribution . If we try to integrate this pdf (probability density function) at infinity convergence requires $\alpha>0$ and the origin it requires $\alpha<-1$ These two conditions can not be simutaneously true and this is why the Pareto distribution is defined in terms of the lower limit $Xm$. Moreover, for the moments to be well defined $m < \alpha$, \cite{WikipediaPowerLaw} which means that that for the mean to be well defined $\alpha>1$ In general the moments of the Pareto distribution are expressed as: \begin{equation} \langle x^{m} \rangle = \int_{xm}^{\infty} x^{m} p(x) dx = { \alpha \over \alpha - m } \left(x_m\right)^m, \; where \;\; m>\alpha \end{equation} From which one can derive: \begin{equation} {\displaystyle \operatorname {E} (X)={\begin{cases}\infty &\alpha \leq 1,\\{\frac {\alpha x_{\mathrm {m} }}{\alpha -1}}&\alpha >1.\end{cases}}} \end{equation} and \begin{equation} {\displaystyle \operatorname {Var} (X)={\begin{cases}\infty &\alpha \in (1,2],\\\left({\frac {x_{\mathrm {m} }}{\alpha -1}}\right)^{2}{\frac {\alpha }{\alpha -2}}&\alpha >2.\end{cases}}} \end{equation} \subsection{Typical values of alpha and the 80-20 rule} Recall from the previous section for the Parato distribution to even converge $\alpha>1$ and for the mean to be well defined $\alpha>2$. Without some further limiting factor (e.g. exponential limiting), the Pareto will always be a bit tail heavy becasue there will be some limit on how high an order of central moments is defined. Furtermore, not only are the higher order moments not gaurnteed to be defined but even the first order moment is heavilyt tail dependent. Sometimes this is known as (Breaking the curve) where a few exceptionaly high values can play a large role in the mean. The rule of thumb for a Pareto distirubtion is that 20\% of all people receive 80\% of all income. As given on wikipedia with this rule we have, \begin{equation} {\displaystyle \alpha_{(80-20)} =\log _{4}5={\cfrac {\log _{10}5}{\log _{10}4}}\approx 1.161} \end{equation} as we will show later this number is close to what one might infer from Oxfam data for wealth. \subsection{Kurtosis} Kurtosis plays an important role in determinging how quickly various estimates of central moments converage and is also a measure of how tail heavy a function is. The Kurtosis for the Pareto distribution is (\href{https://en.wikipedia.org/w/index.php?title=Pareto_distribution&oldid=965348211#Relation_to_the_\%22Pareto_principle\%22}{from wikipedia} \cite{WikipediaParetoDistribution} ): \begin{equation} \text{Excess kurtosis}=\frac{6(\alpha^3+\alpha^2-6\alpha-2)}{\alpha(\alpha-3)(\alpha-4)}\text{ for }\alpha>4 \end{equation} The Pareto distribution has an Excess Kurtosis value which is greater than one. This type of distirubtion is refered to as \href{https://en.wikipedia.org/w/index.php?title=Kurtosis&oldid=965968832#Leptokurtic}{Leptokurtic} \cite{WikipediaKurtosis} and is characterized by a fatter tail. Other examples of such distibutions are Student's t-distribution, Rayleigh distribution, Laplace distribution, Poisson distribution and the logistic distribution. \section{Lorenz curve} \label{LorenzCurve} The Lorenz Curve provides a good way to visualize inequality (\href{https://www.facebook.com/groups/280538506628903/permalink/285154029500684/}{faceook})(\href{https://en.wikipedia.org/w/index.php?title=Pareto_distribution&oldid=967068451#Lorenz_curve_and_Gini_coefficient}{wikipedia} \cite{WikipediaParetoDistribution} \cite{WikipediaLorenzCurve}). \includegraphics[width=0.9\textwidth]{ParetoLorenzSVG.png} The formal definition is: \begin{equation} L(F)=\frac{\int_{x_\mathrm{m}}^{x(F)}xf(x)\,dx}{\int_{x_\mathrm{m}}^\infty xf(x)\,dx} =\frac{\int_0^F x(F')\,dF'}{\int_0^1 x(F')\,dF'} \end{equation} where x(F) is the inverse of the CDF. The CDF is given in euation \eqref{eq:CDF_pareto} and has the following inverse. \begin{equation} x(F)=\frac{x_\mathrm{m}}{(1-F)^{\frac{1}{\alpha}}} \label{eq:inv_cdf_pareto} \end{equation} where, $F(X)=P(xx)=1-F(x) = \left[1+{\frac {x-\mu }{\sigma }}\right]^{-\alpha }} \end{equation} the lomax distribution was refered to by Johnson \& Kotz (1970) \cite{Johnson1970} as a Pareto distribution of the second kind ( \href{https://www.facebook.com/groups/280538506628903/permalink/281081976574556/}{facebook} \cite{Clark1999}), which has the following probability mass function. \begin{equation} {\displaystyle {\displaystyle p(x)={{\alpha \lambda ^{\alpha }} \over {(x+\lambda )^{\alpha +1}}}}={\alpha \over \lambda }\left[{1+{x \over \lambda }}\right]^{-(\alpha +1)},\qquad x\geq 0,} \end{equation} the mean for the Type II pareto distirbution is given by: \begin{equation} E[X]=\frac{ \sigma }{\alpha-1} \end{equation} and in general the central moments are: \begin{equation} E[X^\delta]= \frac{ \sigma^\delta \Gamma(\alpha-\delta)\Gamma(1+\delta)}{\Gamma(\alpha)} \end{equation} where for positive integers (\href{https://en.wikipedia.org/w/index.php?title=Gamma_function&oldid=962235242}{wikipedia} \cite{WikipediaGammaFn}) \begin{equation} {\displaystyle \Gamma (n)=(n-1)!\ .} \end{equation} \subsubsection{Type III \& IV Parato Distributions} A type IV Parato Distriubtion can genearlizes Types I through to III as follows: \begin{equation} P(IV)(\sigma, \sigma, 1, \alpha) = P(I)(\sigma, \alpha), \end{equation} \begin{equation} P(IV)(\mu, \sigma, 1, \alpha) = P(II)(\mu, \sigma, \alpha), \end{equation} \begin{equation} P(IV)(\mu, \sigma, \gamma, 1) = P(III)(\mu, \sigma, \gamma). \end{equation} The survival function for the Type IV pareto distibution is: \begin{equation} {\displaystyle {\overline {F}}(x)=P(X>x)=1-F(x) = \left[1+\left({\frac {x-\mu }{\sigma }}\right)^{1/\gamma }\right]^{-\alpha }} \end{equation} where, ${\displaystyle x\geq \mu }$ and $\mu \in R \;\;$ $\sigma, \gamma > 0, \alpha$ \newline and has the following central moments \begin{equation} E[X^\delta]= \frac{\sigma^\delta\Gamma(\alpha-\gamma \delta)\Gamma(1+\gamma \delta)}{\Gamma(\alpha)} \end{equation} where $\alpha$ is the tail index, $\mu$ is location, $\sigma$ is scale, $\gamma$ is an inequality parameter. \subsection{The Log-Logistic Distirbution} \label{LogLogisticDist} The cumulative distribution for the Type IV pareto distribution can be written as: \begin{equation} {\displaystyle F(x)=1-{\overline F}(x)=P(X1, \; b=\pi /\beta \end{equation} \begin{equation} \operatorname {Var}(X)=\alpha ^{2}\left(2b/\sin 2b-b^{2}/\sin ^{2}b\right),\quad \beta >2, \; b=\pi /\beta \end{equation} \section{Derivation of Log-Type Distributions} The log-logistic distribution is the probability distribution of a random variable whose logarithm has a logistic distribution. In general we can consider a random variable of the form: \begin{equation} X=e^{\mu +\sigma Z} \end{equation} Where Z is a random variable of a given type (e.g. logistic or normal) and X is a variable who is a distribution of that type. Stated formally: \begin{equation} {\displaystyle \ln(X)\sim {\mathcal {N}}(\mu ,\sigma ^{2}).} \end{equation} and it follows: \begin{equation} {\displaystyle {\begin{aligned}f_{X}(x)&={\frac {\rm {d}}{{\rm {d}}x}}\Pr(X\leq x)={\frac {\rm {d}}{{\rm {d}}x}}\Pr(\ln X\leq \ln x)={\frac {\rm {d}}{{\rm {d}}x}}\Phi \left({\frac {\ln x-\mu }{\sigma }}\right)\\[6pt]&=\varphi \left({\frac {\ln x-\mu }{\sigma }}\right){\frac {\rm {d}}{{\rm {d}}x}}\left({\frac {\ln x-\mu }{\sigma }}\right)=\varphi \left({\frac {\ln x-\mu }{\sigma }}\right){\frac {1}{\sigma x}}\\[6pt].\end{aligned}}} \end{equation} and if $\phi(x)$ is a normal distrubution then \begin{equation} f_{X}(x)={\frac {1}{x}}\cdot {\frac {1}{\sigma {\sqrt {2\pi \,}}}}\exp \left(-{\frac {(\ln x-\mu )^{2}}{2\sigma ^{2}}}\right) \end{equation} alternatively if $\phi(x)$ is a logistic distrubution then \begin{equation} f(x;\alpha ,\beta )={\frac {(\beta /\alpha )(x/\alpha )^{{\beta -1}}}{\left(1+(x/\alpha )^{{\beta }}\right)^{2}}} \end{equation} \section{The Quantile Function} In section \ref{LogLogisticDist} we used the aproximation $\alpha=1$ to argue that the Pareto distribution reduces to the log-logistic distirubtion when $\mu=0$. As stated, this aproximation is exact for a "\emph{Type III pareto distribution}". The "\emph{Type III Pareto Distribugion}", generalizes the \emph{"Type I Pareto Distirubution"} and all types (i.e.Types I to IV) converge to the "\emph{Type I Pareto Distribution}" for large values of $X$. For the following analysis with the cumulative distirbution function given in equation \eqref{eq:GenLogLogDist} since it is more general than the log-logistic distirubtion. We can invert this equation by solving for $X$ in terms of the value of the cumulative distribution function. \begin{equation} x=\mu + \left[ {\sigma^\beta F(x) \over 1 - F(x) } \right]^{\left( 1/\beta\right) }=\mu + \sigma \left[ {F(x) \over 1 - F(x) } \right]^{\left( 1/\beta\right) } \end{equation} \label{eq:inv_cdf_paretoIII} The result is the Quantile Function for the Type III Pareto distibution and if we set $\mu=0$ this is the Quantile FUnction for the log-logistic distirbution (\href{https://en.wikipedia.org/w/index.php?title=Logistic_distribution&oldid=955853904#Quantile_function}{wikipedia} \cite{WikipediaQuantileFunction} )(\href{https://www.facebook.com/groups/280538506628903/permalink/280655966617157/}{facebook}). \subsection{The Asymptoic Quantile Function For Types I \& III Parato Distributions} The form of the Type III Pareto Distirugiton Quantile Function \eqref{eq:inv_cdf_paretoIII} is notacibly different than the Type I Parto Distirguion Quantile Function (i.e. equation \eqref{eq:inv_cdf_pareto}). The main distinquishing factor is the $F(x)$ in the numberator is not present in the Type I version of the Quantile Function. The asymptoic simmilarity can be shown by using a new variable $\epsilon = 1-F(x)$. With this substitution \eqref{eq:inv_cdf_paretoIII} becomes: \begin{equation} x=\mu + \sigma \left[ \frac{1}{\epsilon} -1 \right]^{\left( 1/\beta\right) }\cong \mu + \sigma\left[ \frac{1}{\epsilon} \right]^{\left( 1/\beta\right) } \end{equation} \label{eq:inv_cdf_paretoIII} when both $\mu=0$ and $\sigma=(x_m)^{1/\beta}$ we get the same asymptotic result for both the types I and III pareto distributions. \section{The Isograph} The Isograph can be obtained by first subtracing $\mu$ from each side of equation \eqref{eq:inv_cdf_paretoIII} and then dividing by the median to yield a log-logit transformation of the Type III pareto disribution. This should be s stright line with slope $(1/\beta)$ and intercept $\frac{\sigma}{\sigma_{median}}=\sigma^\prime$ \newline As can be scene this log-logit transformation produces a faily straight line \cite{isographslide} \newline \includegraphics[width=1.2\textwidth]{LogitWealth.png} Dividing the Y variable \cite{Mishra2017} \cite{Chauvel2018} we get the isograph, which is a constant when the curve is a Type III pareto distribution. The deviation from this constant represents unexpected inequality within a given group. Typically the logrithm used for this graph is the natural logarithm. The isograph essentially shows where the Type III pareto is not a good fit. \section{The Quantile Function} \begin{thebibliography}{9} \bibitem{wikipediaSurvivalFunction} Wikipedia, Survival function, \url{https://en.wikipedia.org/wiki/Survival_function} (\href{https://www.facebook.com/groups/280538506628903/permalink/281646936518060/}{Facebook:281646936518060}) \begin{verbatim} Breadcrumb: https://www.facebook.com/groups/280538506628903/permalink/281646936518060/\end{verbatim} \bibitem{JulietteFournierSep2015} JulietteFournier Sep 2015, Generalized Pareto curves:Theory and application using income and inheritancetabulations for France 1901-2012, \url{http://piketty.pse.ens.fr/files/Fournier2015.pdf} (\href{https://www.facebook.com/groups/280538506628903/permalink/281658899850197/}{Facebook:281658899850197}) \bibitem{Persky1992} Persky, J. (1992). Retrospectives: Pareto’s Law.Journal of Economic Perspectives, 6(2):181–192. \bibitem{WikipediaPowerLaw} \url{https://en.wikipedia.org/w/index.php?title=Power_law&oldid=965150788#Power-law_probability_distributions} \bibitem{WikipediaParetoDistribution} \url{https://en.wikipedia.org/w/index.php?title=Pareto_distribution&oldid=965348211#Relation_to_the_\%22Pareto_principle\%22} \bibitem{WikipediaKurtosis} \url{https://en.wikipedia.org/w/index.php?title=Kurtosis&oldid=965968832#Leptokurtic} \bibitem{WikipediaParetoPrinciple} \url{https://en.wikipedia.org/w/index.php?title=Pareto_principle&oldid=964003639} \bibitem{WikipediaLorenzCurve} \url{https://en.wikipedia.org/w/index.php?title=Lorenz_curve&oldid=946106761} \bibitem{WikipediaQuantileFunction} \url{https://en.wikipedia.org/w/index.php?title=Logistic_distribution&oldid=955853904#Quantile_function} \bibitem{WikipediaLomax} \url{https://en.wikipedia.org/w/index.php?title=Lomax_distribution&oldid=958008294#Characterization} \bibitem{Clark1999} R. M. Clark (1999), Generalizations of power-law distributions applicable to sampledfault-trace lengths: model choice, parameter estimation and caveats, Geophys. J. Int.(1999)136,357\^372 \bibitem{Johnson1970} Johnson, N.L. and Kotz (1970). Continuous Univariate Distributions I, Houghton Mi, New York \bibitem{WikipedaGammaFunction} Aswini Kumar Mishra, EXAMINING CHANGES IN THE LEVEL AND SHAPE OF INCOME DISTRIBUTIONS IN INDIA, 2005-2012 \url{https://en.wikipedia.org/w/index.php?title=Gamma_function&oldid=962235242} \bibitem{isographslide} Slide 18 of \url{https://slideplayer.com/slide/14345914/}, Income and wealth distribution \bibitem{Mishra2017} Aswini Kumar Mishra, EXAMINING CHANGES IN THE LEVEL AND SHAPE OF INCOME DISTRIBUTIONS IN INDIA, 2005-2012 pg 5 of \url{https://pdfs.semanticscholar.org/58c2/c262230b2a150a1d4c68d1806d9cbe5e7212.pdf} for isograph deffintion \bibitem{Chauvel2018} Louis Chauvel (2018), Wealth as an Increasing Source of Inequality and Distortion in Income Groups pg 5 of \url{http://www.iariw.org/copenhagen/chauvel.pdf} for isograph deffintion \bibitem{einstein} Albert Einstein. \textit{Zur Elektrodynamik bewegter K{\"o}rper}. (German) [\textit{On the electrodynamics of moving bodies}]. Annalen der Physik, 322(10):891–921, 1905. \bibitem{knuthwebsite} Knuth: Computers and Typesetting, \\\texttt{http://www-cs-faculty.stanford.edu/\~{}uno/abcde.html} \end{thebibliography} \end{document}