Essay

Artificial Intelligence (Machine and Deep Learning)

Rating

Sold

Pages

Uploaded on

25-11-2024

Written in

2023/2024

This documents contains assignment questions from forth year Statistic Courses (Specifically Artificial Intelligent): 1. The Newton-Raphson Method. 2. Iteratively Reweighted Least Squares (IRLS). 3. Survival Analysis (hazard function H(t), Survival function S(t), probability density function f(t)). 4. Proving the equivalence between Bayes Rule and Minimum Expected Loss Rule for classification under the assumption of known class-conditional densities. 5. Performing the back propagation algorithm step-by-step for a single two-layer neural network with Sigmoid activation function. Derive the update equations for the weights and biases. 6. Formalizing the Universal Approximation Theorem for Neural Networks and explain its implications for the expressive power of these models. Discuss the limitations of the theorem and it's relationship to the number of neurons and network architecture. 7. Analysing the Vapnik-Chervonenkis (VC) dimension of linear regression and explain how it relates to the generalization ability of the model. 8. Formulating the decision tree learning problem as a constrained optimization problem with the objective of minimising the expected risk and constraints based on tree depth and complexity. Discuss the trade-off between these factors.

Show more Read less

Institution

Course

Content preview

The Significance or Role of the Parameter 𝜽 in Statistical Theory and Reality

To begin with, capital letter Θ is used to denote the parameter space: all possible values the
variable can potentially take. It is also used with an index such as Θ0 to denote the space under
the null hypothesis and Θ1 to denote the space under the alternative hypothesis.

However, the Greek small letter 𝜃 is used in statistic to denote an unknown parameter of interest.
For an example in A/B testing it is usually modeled as a random variable. The true value of 𝜃 is
denoted by 𝜃 ∗ , while the estimator of 𝜃 (usually the likelihood estimate) is denoted with a hat
above the letter.

The common problem is to find the values of 𝜃. For example: it is commonly used to name the
mean and the standard deviation in a Normal distribution mu and sigma 𝑁 ∼ (𝜇, 𝜎). 𝜇 in Normal
distribution tells where the mean of the distribution is and so it can describe random variables
with different mean values. So, the parameters are often called 𝜃. Other distributions have at
least one unknown parameter of interest. In Binomial distribution, there are two parameters: the
number of independent trials (n) and the probability of success (p). There is also the Gamma
distribution consisting of two parameters: the shape parameter 𝛼 and the rate parameter 𝛽.
Furthermore, the is the Poisson distribution consisting of only one parameter which is the mean
number of events 𝜆. There is also the Geometric distribution which has only one parameter,
which is the probability of success for each trial (p).

Theta 𝜃 can apply in any parameters you want to estimate. For an example: the reference of 𝜃
can depend on what model you are working on, such as the least squares regression where you
model a dependent variable (Y) as a linear combination of one or more independent variables
(X). 𝑌𝑖 = 𝑏0 + 𝑏1 𝑋1 + 𝑏2 𝑋2 + ⋯ + 𝑏𝑛 𝑋𝑛 . Where n is the number of independent variables and
the parameters to be estimated are the 𝛽𝑠 . So, 𝜃 is the name of all the 𝛽𝑠 .

On another example: you want to study the disintegration of radioactive atom which decreases
exponentially. Letting t to be the time to disintegration then the model is: 𝑓(𝑡) = 𝜃𝑒 −𝜃𝑡 . Where
f(t) is a probability density function of an atom disintegrating in the time interval (t, t + dt) which
is f(t) dt. So, the interest is to estimate 𝜃 which is the disintegration rate.

On the last example: You want to study the precision of a weighing instrument. (measurements
are Gaussian), you model the weighing of a standard 1kg object as: 𝑓(𝑥) =
1 𝑥−𝜇 2
exp {− ( 2𝜎 ) } . Where x is the measure given by the scale. F(x) is the probability density
𝜎√2𝜋
and the parameters are 𝜇 𝑎𝑛𝑑 𝜎. So, 𝜃 = (𝜇, 𝜎) where mu is the target weight and sigma is the
standard deviation of the measure every time you weigh the object.

In conclusion, symbol 𝜃 is used to denote any unknown parameters of interest. Statistics is about
finding the best or appropriate 𝜃 values (Bayesians would say: given the data and priors.

Report Copyright Violation

Written for

Course: STA

All documents for this subject (54)

Document information

Uploaded on: November 25, 2024
Number of pages: 3
Written in: 2023/2024
Type: ESSAY
Professor(s): Unknown
Grade: Unknown

Subjects

newton raphson method
survival analysis
bayes rule
back propagation for n
vapnik chervonenkis
universal
iteratively reweighted least squares
probability density function
decision tree learning problem

$21.16

Get access to the full document:

Written by students who passed

Immediately available after payment

Read online or as PDF

Get to know the seller

phephymalinga58

Get to know the seller

phephymalinga58 University of Swaziland

View profile

Sold

Member since

1 year

Number of followers

Documents

Last sold

0.0

0 reviews

Why students choose Stuvia

Created by fellow students, verified by reviews

Quality you can trust: written by students who passed their tests and reviewed by others who've used these notes.

Didn't get what you expected? Choose another document

No worries! You can instantly pick a different document that better fits what you're looking for.

Pay as you like, start learning right away

No subscription, no commitments. Pay the way you're used to via credit card and download your PDF document instantly.

“Bought, downloaded, and aced it. It really can be that simple.”

Alisha Student

Frequently asked questions

What do I get when I buy this document?

You get a PDF, available immediately after your purchase. The purchased document is accessible anytime, anywhere and indefinitely through your profile.

Satisfaction guarantee: how does it work?

Our satisfaction guarantee ensures that you always find a study document that suits you well. You fill out a form, and our customer service team takes care of the rest.

Who am I buying these notes from?

Stuvia is a marketplace, so you are not buying this document from us, but from seller phephymalinga58. Stuvia facilitates payment to the seller.

Will I be stuck with a subscription?

No, you only buy these notes for $21.16. You're not tied to anything after your purchase.

Can Stuvia be trusted?

4.6 stars on Google & Trustpilot (+1000 reviews) 47251 documents were sold in the last 30 days Founded in 2010, the go-to place to buy study notes for 16 years now

Artificial Intelligence (Machine and Deep Learning)

Content preview

Written for

Document information

Subjects

Get to know the seller

Recently viewed by you

Why students choose Stuvia

Created by fellow students, verified by reviews

Didn't get what you expected? Choose another document

Pay as you like, start learning right away

Working on your references?

Frequently asked questions

What do I get when I buy this document?

Satisfaction guarantee: how does it work?

Who am I buying these notes from?

Will I be stuck with a subscription?

Can Stuvia be trusted?