Maximum Likelihood及Maximum Likelihood Estimation
1、What is Maximum Likelihood?
极大似然是一种找到最可能解释一组观测数据的函数的方法。
Maximum Likelihood is a way to find the most likely function to explain a set of observed data.
在基本统计学中,通常给你一个模型来计算概率。例如,你可能被要求找出X大于2的概率,给定如下泊松分布:X ~ Poisson (2.4)。在这个例子中,已经给定了你泊松分布的参数 λ(2.4),在现实生活中,您没有这么奢侈,因为您没有确定参数的模型:您必须将数据与模型相匹配。这就是最大可能性(MLE)的作用。在统计学中,最大似然估计(maximum likelihood estimation, MLE)是在给定观测值的情况下估计统计模型参数的一种方法。MLE试图在给定观测值的情况下找到使似然函数最大化的参数值。得到的估计称为最大似然估计,也缩写为MLE。
In elementary statistics, you are usually given a model to find probabilities. For example, you might be asked to find the probability that X is greater than 2, given the following Poisson distribution:
X ~ Poisson (2.4)
In this example, you are given the parameter, λ, of 2.4 for the Possion distribution. In real life, you don’t have the luxury of having a model given to you: you’ll have to fit your data to a model. That’s where Maximum Likelihood (MLE) comes in.
In statistics, maximum likelihood estimation (MLE) is a method of estimating the parameters of a statistical model, given observations. MLE attempts to find the parameter values that maximize the likelihood function, given the observations. The resulting estimate is called a maximum likelihood estimate, which is also abbreviated as MLE.
MLE采用已知的概率分布模型(如正态分布),并将数据集与这些分布进行比较,以便找到数据的合适匹配。一个分布模型对应的参数可以有无穷个。例如正态分布的均值可以是0,也可以是100亿以上。最大似然估计是找到最可能生成待测样本的总体参数的一种方法。数据与模型的匹配程度称为“拟合优度”。
MLE takes known probability distributions (like the normal distribution) and compares data sets to those distributions in order to find a suitable match for the data. A Family of distributions can have an infinite amount of possible parameters. For example, the mean of the normal distribution could be equal to zero, or it could be equal to ten billion and beyond. Maximum Likelihood Estimation is one way to find the parameters of the population that is most likely to have generated the sample being tested. How well the data matches the model is known as “Goodness of Fit.”
例如,研究人员可能有兴趣找出吃特定食物的老鼠的平均体重增加。研究人员无法测量每只老鼠的体重,所以只能取样。大鼠体重增加呈正态分布;最大似然估计可用于求基于该样本的总体增重的均值和方差
For example, a researcher might be interested in finding out the mean weight gain of rats eating a particular diet. The researcher is unable to weigh every rat in the population so instead takes a sample. Weight gains of rats tend to follow a normal distribution; Maximum Likelihood Estimation can be used to find the mean and variance of the weight gain in the general population based on this sample
MLE根据似然函数的最大值来选择模型参数。
MLE chooses the model parameters based on the values that maximize the Likelihood Function.
2、The Likelihood Function(似然函数,是一种表示概率的方法;似然表示得到样本的概率;最大似然表示的是得到样本最大概率的参数)
给定一个特定的概率分布模型,样本的似然是得到样本的概率。似然函数是一种表示概率的方法:最大概率得到样本的参数是最大似然估计。
一句话:似然表示概率;似然函数表示得到概率的方法;最大似然表示的得到最大概率的参数
The likelihood of a sample is the probability of getting that sample, given a specified probability distribution model. The likelihood function is a way to express that probability: the parameters that maximize the probability of getting that sample are the Maximum Likelihood Estimators.
假设你有一组从一个未知分布参数Θ的总体得到的随机变量X1, X2…Xn。该分布的概率密度函数(PDF) f(Xi,Θ)模型,Xi是随机变量的集合,Θ是未知参数。最大似然函数你想知道Θ最可能的值是什么,得到随机变量Xi。本例的联合概率密度函数为:
Let’s suppose you had a set of random variables X1, X2…Xn taken from an unknown population distribution with parameter Θ. This distribution has a probability density function (PDF) of f(Xi,Θ) where f is the model, Xi is the set of random variables and Θ is the unknown parameter. For the maximum likelihood function you want to know what the most likely value for Θ is, given the set of random variables Xi. The joint probability density function for this example is:

3、The Basic Idea
It seems reasonable that a good estimate of the unknown parameter θ would be the value of θ that maximizes the probability, errrr... that is, the likelihood... of getting the data we observed. (So, do you see from where the name "maximum likelihood" comes?) So, that is, in a nutshell, the idea behind the method of maximum likelihood estimation. But how would we implement the method in practice? Well, suppose we have a random sample X1, X2,..., Xn for which the probability density (or mass) function of each Xi is f(xi; θ). Then, the joint probability mass (or density) function of X1, X2,..., Xn, which we'll (not so arbitrarily) call L(θ) is:

The first equality is of course just the definition of the joint probability mass function. The second equality comes from that fact that we have a random sample, which implies by definition that the Xi are independent. And, the last equality just uses the shorthand mathematical notation of a product of indexed terms. Now, in light of the basic idea of maximum likelihood estimation, one reasonable way to proceed is to treat the "likelihood function" L(θ) as a function of θ, and find the value of θ that maximizes it.
4、example1
假设权重随机选择的美国女大学生与未知的正态分布均值μ和标准差σ。随机抽取的10名美国女大学生的体重(以磅为单位)如下:
115 122 130 127 149 160 152 138 149 180
根据上面给出的定义,识别似然函数和μ的极大似然估计量,所有的美国女大学生的平均重量。使用给定的样本,找到一个最大似然估计的μ。
Based on the definitions given above, identify the likelihood function and the maximum likelihood estimator of μ, the mean weight of all American female college students. Using the given sample, find a maximum likelihood estimate of μ as well.
5、example2
Suppose we have a random sample X1, X2,..., Xn where:
- Xi = 0 if a randomly selected student does not own a sports car, and
- Xi = 1 if a randomly selected student does own a sports car.
Assuming that the Xi are independent Bernoulli random variables with unknown parameter p, find the maximum likelihood estimator of p, the proportion of students who own a sports car.




6、文献
https://newonlinecourses.science.psu.edu/stat414/node/191/(写的很好,里面有很多的例子)
https://en.wikipedia.org/wiki/Maximum_likelihood_estimation
https://www.statisticshowto.datasciencecentral.com/maximum-likelihood-estimation/
Maximum Likelihood及Maximum Likelihood Estimation的更多相关文章
- MLE vs MAP: the connection between Maximum Likelihood and Maximum A Posteriori Estimation
Reference:MLE vs MAP. Maximum Likelihood Estimation (MLE) and Maximum A Posteriori (MAP), are both a ...
- LeetCode: Maximum Product Subarray && Maximum Subarray &子序列相关
Maximum Product Subarray Title: Find the contiguous subarray within an array (containing at least on ...
- likelihood(似然) and likelihood function(似然函数)
知乎上关于似然的一个问题:https://www.zhihu.com/question/54082000 概率(密度)表达给定下样本随机向量的可能性,而似然表达了给定样本下参数(相对于另外的参数)为真 ...
- [Bayes] Understanding Bayes: A Look at the Likelihood
From: https://alexanderetz.com/2015/04/15/understanding-bayes-a-look-at-the-likelihood/ Reading note ...
- [LeetCode] Maximum Depth of Binary Tree 二叉树的最大深度
Given a binary tree, find its maximum depth. The maximum depth is the number of nodes along the long ...
- LeetCode 104. Maximum Depth of Binary Tree
Problem: Given a binary tree, find its maximum depth. The maximum depth is the number of nodes along ...
- [LintCode] Maximum Depth of Binary Tree 二叉树的最大深度
Given a binary tree, find its maximum depth. The maximum depth is the number of nodes along the long ...
- [Leetcode][JAVA] Minimum Depth of Binary Tree && Balanced Binary Tree && Maximum Depth of Binary Tree
Minimum Depth of Binary Tree Given a binary tree, find its minimum depth. The minimum depth is the n ...
- LeetCode:Maximum Depth of Binary Tree_104
LeetCode:Maximum Depth of Binary Tree [问题再现] Given a binary tree, find its maximum depth. The maximu ...
随机推荐
- C#打印0到100的素数
static void Main(string[] args) { //输出1-100的素数 bool res; ; ; i < ; i++) { res = true; ; j < i; ...
- IdentityServer4 接口说明
在.net core出来以后很多人使用identityServer做身份验证. ids4和ids3的token验证组件都是基于微软的oauth2和bearer验证组件.园子里也很多教程,我们通过教程了 ...
- ssh Socket error Event: 32 Error: 10053.
在家用的WiFi,把电脑从房间搬到餐厅来用发现用我的xshell不能用ssh连接了,报错Socket error Event: 32 Error: 10053.同时在自己物理机上ipconfig看到自 ...
- 《算法》第六章部分程序 part 4
▶ 书中第六章部分程序,包括在加上自己补充的代码,利用后缀树查找最长重复子串.查找最大重复子串并输出其上下文(Key word in context,KWIC).求两字符串的最长公共子串 ● 利用后缀 ...
- 配置Eclipse的Maven环境
- java -version 问题 : C:\ProgramData\Oracle\Java\javapath;
我把 JAVA_HOME 从8改成了 7 , 为什么还是 显示的8啊 ! E:\sv0\jars>java -version java version "1.8.0_111" ...
- 【ASP.NET 插件】分享一款富文本web编辑器UEditor
UEditor是由百度web前端研发部开发所见即所得富文本web编辑器,具有轻量,可定制,注重用户体验等特点,开源基于MIT协议,允许自由使用和修改代码... <%@ Page Language ...
- ssh-keygen生成git ssh密钥
title: ssh-keygen生成git ssh密钥 date: 2018-05-07 08:49:21 tags: [git,ssh-keygen] --- ssh-keygen生成git ss ...
- python_04 基本数据类型、数字、字符串、列表、元组、字典
基本数据类型 所有的方法(函数)都带括号,且括号内没带等号的参数需传给它一个值,带等号的参数相当于有默认值 1.数字 int 在32位机器上,整数的位数为32位,取值范围为-2**31-2**31-1 ...
- django 认证模块auth,表单组件form
django认证系统(auth): 1.首先我们在新窗口中打开一个django项目,之后点击,
