贝叶斯统计模型java代码_Pymc - 贝叶斯统计模型

Introduction

Version:

2.3.6

Authors:

Chris Fonnesbeck

Anand Patil

David Huard

John Salvatier

Copyright:

This document has been placed in the public domain.

License:PyMC is released under the Academic Free License.

9087c0e24bc9ff4462cfffd009d1ab2f.png

687474703a2f2f696d672e736869656c64732e696f2f707970692f762f70796d632e7376673f7374796c653d666c6174

687474703a2f2f696d672e736869656c64732e696f2f62616467652f6c6963656e73652d41464c2d626c75652e7376673f7374796c653d666c6174

NOTE: The development version of PyMC (version 3) has been moved to its own repository called pymc3.

Purpose

PyMC is a python module that implements Bayesian statistical models and

fitting algorithms, including Markov chain Monte Carlo.

Its flexibility and extensibility make it applicable to a large suite of problems. Along with core sampling functionality, PyMC includes

methods for summarizing output, plotting, goodness-of-fit and convergence

diagnostics.

Features

PyMC provides functionalities to make Bayesian analysis as painless as

possible. Here is a short list of some of its features:

Fits Bayesian statistical models with Markov chain Monte Carlo and

other algorithms.

Includes a large suite of well-documented statistical distributions.

Uses NumPy for numerics wherever possible.

Includes a module for modeling Gaussian processes.

Sampling loops can be paused and tuned manually, or saved and restarted later.

Creates summaries including tables and plots.

Traces can be saved to the disk as plain text, Python pickles, SQLite or MySQL

database, or hdf5 archives.

Several convergence diagnostics are available.

Extensible: easily incorporates custom step methods and unusual probability

distributions.

MCMC loops can be embedded in larger programs, and results can be analyzed

with the full power of Python.

What's new in version 2

This second version of PyMC benefits from a major rewrite effort.

Substantial improvements in code extensibility, user interface as well

as in raw performance have been achieved. Most notably, the PyMC 2 series

provides:

New flexible object model and syntax (not backward-compatible).

Reduced redundant computations: only relevant log-probability terms are

computed, and these are cached.

Optimized probability distributions.

New adaptive blocked Metropolis step method.

Much more!

Usage

First, define your model in a file, say mymodel.py (with comments, of course!):

# Import relevant modules

import pymc

import numpy as np

# Some data

n = 5 * np.ones(4, dtype=int)

x = np.array([-.86, -.3, -.05, .73])

# Priors on unknown parameters

alpha = pymc.Normal('alpha', mu=0, tau=.01)

beta = pymc.Normal('beta', mu=0, tau=.01)

# Arbitrary deterministic function of parameters

@pymc.deterministic

def theta(a=alpha, b=beta):

"""theta = logit^{-1}(a+b)"""

return pymc.invlogit(a + b * x)

# Binomial likelihood for data

d = pymc.Binomial('d', n=n, p=theta, value=np.array([0., 1., 3., 5.]),

observed=True)复制代码

Save this file, then from a python shell (or another file in the same directory), call:

import pymc

import mymodel

S = pymc.MCMC(mymodel, db='pickle')

S.sample(iter=10000, burn=5000, thin=2)

pymc.Matplot.plot(S)复制代码

This example will generate 10000 posterior samples, thinned by a factor of 2, with the first half discarded as burn-in. The sample is stored in a Python serialization (pickle) database.

History

PyMC began development in 2003, as an effort to generalize the process of building Metropolis-Hastings samplers, with an aim to making Markov chain Monte Carlo (MCMC) more accessible to non-statisticians (particularly ecologists). The choice to develop PyMC as a python module, rather than a standalone application, allowed the use MCMC methods in a larger modeling framework. By 2005, PyMC was reliable enough for version 1.0 to be released to the public. A small group of regular users, most associated with the University of Georgia, provided much of the feedback necessary for the refinement of PyMC to a usable state.

In 2006, David Huard and Anand Patil joined Chris Fonnesbeck on the development team for PyMC 2.0. This iteration of the software strives for more flexibility, better performance and a better end-user experience than any previous version of PyMC.

PyMC 2.1 was released in early 2010. It contains numerous bugfixes and optimizations, as well as a few new features. This user guide is written for version 2.1.

Relationship to other packages

PyMC in one of many general-purpose MCMC packages. The most prominent among them is WinBUGS, which has made MCMC and with it Bayesian statistics accessible to a huge user community. Unlike PyMC, WinBUGS is a stand-alone, self-contained application. This can be an attractive feature for users without much programming experience, but others may find it constraining. A related package is JAGS, which provides a more UNIX-like implementation of the BUGS language. Other packages include Hierarchical Bayes Compiler and a number of R packages of varying scope.

It would be difficult to meaningfully benchmark PyMC against these other packages because of the unlimited variety in Bayesian probability models and flavors of the MCMC algorithm. However, it is possible to anticipate how it will perform in broad terms.

PyMC's number-crunching is done using a combination of industry-standard libraries (NumPy and the linear algebra libraries on which it depends) and hand-optimized Fortran routines. For models that are composed of variables valued as large arrays, PyMC will spend most of its time in these fast routines. In that case, it will be roughly as fast as packages written entirely in C and faster than WinBUGS. For finer-grained models containing mostly scalar variables, it will spend most of its time in coordinating Python code. In that case, despite our best efforts at optimization, PyMC will be significantly slower than packages written in C and on par with or slower than WinBUGS. However, as fine-grained models are often small and simple, the total time required for sampling is often quite reasonable despite this poorer performance.

We have chosen to spend time developing PyMC rather than using an existing package primarily because it allows us to build and efficiently fit any model we like within a full-fledged Python environment. We have emphasized extensibility throughout PyMC's design, so if it doesn't meet your needs out of the box chances are you can make it do so with a relatively small amount of code. See the testimonials page on the wiki for reasons why other users have chosen PyMC.

Getting started

This guide provides all the information needed to install PyMC, code

a Bayesian statistical model, run the sampler, save and visualize the results.

In addition, it contains a list of the statistical distributions currently available. More examples of usage as well as

tutorials are available from the PyMC web site.

评论
添加红包

请填写红包祝福语或标题

红包个数最小为10个

红包金额最低5元

当前余额3.43前往充值 >
需支付:10.00
成就一亿技术人!
领取后你会自动成为博主和红包主的粉丝 规则
hope_wisdom
发出的红包
实付
使用余额支付
点击重新获取
扫码支付
钱包余额 0

抵扣说明:

1.余额是钱包充值的虚拟货币,按照1:1的比例进行支付金额的抵扣。
2.余额无法直接购买下载,可以购买VIP、付费专栏及课程。

余额充值