The Mathematical Foundations of Manifold Learning Arxiv
Total Page:16
File Type:pdf, Size:1020Kb
Load more
Recommended publications
-
ECS289: Scalable Machine Learning
ECS289: Scalable Machine Learning Cho-Jui Hsieh UC Davis Nov 3, 2015 Outline PageRank Semi-supervised Learning Label Propagation Manifold regularization PageRank Ranking Websites Text based ranking systems (a dominated approach in the early 90s) Compute the similarity between query and websites (documents) Keywords are a very limited way to express a complex information Need to rank websites by \popularity", \authority", . PageRank: Developed by Brin and Page (1999) Determine the authority and popularity by hyperlinks PageRank Main idea: estimate the ranking of websites by the link structure. Topology of Websites Transform the hyperlinks to a directed graph: The adjacency matrix A such that Aij = 1 if page j points to page i Transition Matrix Normalize the adjacency matrix so that the matrix is a stochastic matrix (each column sum up to 1) Pij : probability that arriving at page i from page j P: a stochastic matrix or a transition matrix Random walk: step 1 Random walk through the transition matrix Start from [1; 0; 0; 0] (can use any initialization) Random walk: step 1 Random walk through the transition matrix xt+1 = Pxt Random walk: step 2 Random walk through the transition matrix Random walk: step 2 Random walk through the transition matrix xt+2 = Pxt+1 PageRank (convergence) P Start from an initial vector x with i xi = 1 (the initial distribution) For t = 1; 2;::: xt+1 = Pxt Each xt is a probability distribution (sums up to 1) Converges to a stationary distribution π such that π = Pπ if P satisfies the following two conditions: t 1 P is irreducible: for all i; j, there exists some t such that (P )i;j > 0 t 2 P is aperiodic: for all i; j, we have gcdft :(P )i;j > 0g = 1 π is the unique right eigenvector of P with eigenvalue 1. -
Semi-Supervised Classification Using Local and Global Regularization
Proceedings of the Twenty-Third AAAI Conference on Artificial Intelligence (2008) Semi- supervised Classification Using Local and Global Regularization Fei Wang1, Tao Li2, Gang Wang3, Changshui Zhang1 1Department of Automation, Tsinghua University, Beijing, China 2School of Computing and Information Sciences, Florida International University, Miami, FL, USA 3Microsoft China Research, Beijing, China Abstract works concentrate on the derivation of different forms of In this paper, we propose a semi-supervised learning (SSL) smoothness regularizers, such as the ones using combinato- algorithm based on local and global regularization. In the lo- rial graph Laplacian (Zhu et al., 2003)(Belkin et al., 2006), cal regularization part, our algorithm constructs a regularized normalized graph Laplacian (Zhou et al., 2004), exponen- classifier for each data point using its neighborhood, while tial/iterative graph Laplacian (Belkin et al., 2004), local lin- the global regularization part adopts a Laplacian regularizer ear regularization (Wang & Zhang, 2006)and local learning to smooth the data labels predicted by those local classifiers. regularization (Wu & Sch¨olkopf, 2007), but rarely touch the We show that some existing SSL algorithms can be derived problem of how to derive a more efficient loss function. from our framework. Finally we present some experimental In this paper, we argue that rather than applying a global results to show the effectiveness of our method. loss function which is based on the construction of a global predictor using the whole data set, it would be more desir- Introduction able to measure such loss locally by building some local pre- Semi-supervised learning (SSL), which aims at learning dictors for different regions of the input data space. -
CMI Summer Schools
CMI Summer Schools 2007 Homogeneous flows, moduli spaces, and arithmetic De Giorgi Center, Pisa 2006 Arithmetic Geometry The Clay Mathematics Institute has Mathematisches Institut, conducted a program of research summer schools Georg-August-Universität, Göttingen since 2000. Designed for graduate students and PhDs within five years of their degree, the aim of 2005 Ricci Flow, 3-manifolds, and Geometry the summer schools is to furnish a new generation MSRI, Berkeley of mathematicians with the knowledge and tools needed to work successfully in an active research CMI summer schools 2004 area. Three introductory courses, each three weeks Floer Homology, Gauge Theory, and in duration, make up the core of a typical summer Low-dimensional Topology school. These are followed by one week of more Rényi Institute, Budapest advanced minicourses and individual talks. Size is limited to roughly 100 participants in order to 2003 Harmonic Analysis, Trace Formula, promote interaction and contact. Venues change and Shimura Varieties from year to year, and have ranged from Cambridge, Fields Institute, Toronto Massachusetts to Pisa, Italy. The lectures from each school are published in the CMI–AMS proceedings 2002 Geometry and String Theory series, usually within two years’ time. Newton Institute, Cambridge UK Summer Schools 2001 www.claymath.org/programs/summer_school Minimal surfaces MSRI, Berkeley Summer School Proceedings www.claymath.org/publications 2000 Mirror Symmetry Pine Manor College, Boston 2006 Arithmetic Geometry Summer School in Göttingen techniques are drawn from the theory of elliptic The 2006 summer school program will introduce curves, including modular curves and their para- participants to modern techniques and outstanding metrizations, Heegner points, and heights. -
Linear Manifold Regularization for Large Scale Semi-Supervised Learning
Linear Manifold Regularization for Large Scale Semi-supervised Learning Vikas Sindhwani [email protected] Partha Niyogi [email protected] Mikhail Belkin [email protected] Department of Computer Science, University of Chicago, Hyde Park, Chicago, IL 60637 Sathiya Keerthi [email protected] Yahoo! Research Labs, 210 S Delacey Avenue, Pasadena, CA 91105 Abstract tion (Belkin, Niyogi & Sindhwani, 2004; Sindhwani, The enormous wealth of unlabeled data in Niyogi and Belkin, 2005). In this framework, unla- many applications of machine learning is be- beled data is incorporated within a geometrically mo- ginning to pose challenges to the designers tivated regularizer. We specialize this framework for of semi-supervised learning methods. We linear classification, and focus on the problem of deal- are interested in developing linear classifica- ing with large amounts of data. tion algorithms to efficiently learn from mas- This paper is organized as follows: In section 2, we sive partially labeled datasets. In this paper, setup the problem of linear semi-supervised learning we propose Linear Laplacian Support Vector and discuss some prior work. The general Manifold Machines and Linear Laplacian Regularized Regularization framework is reviewed in section 3. In Least Squares as promising solutions to this section 4, we discuss our methods. Section 5 outlines problem. some research in progress. 1. Introduction 2. Linear Semi-supervised Learning The problem of linear semi-supervised clasification is A single web-crawl by search engines like Yahoo and setup as follows: We are given l labeled examples Google indexes billions of webpages. Only a very l l+u fxi; yigi=1 and u unlabeled examples fxigi=l+1, where small fraction of these web-pages can be hand-labeled d and assembled into topic directories. -
Kähler-Einstein Metrics on Fano Manifolds. I
JOURNAL OF THE AMERICAN MATHEMATICAL SOCIETY Volume 28, Number 1, January 2015, Pages 183–197 S 0894-0347(2014)00799-2 Article electronically published on March 28, 2014 KAHLER-EINSTEIN¨ METRICS ON FANO MANIFOLDS. I: APPROXIMATION OF METRICS WITH CONE SINGULARITIES XIUXIONG CHEN, SIMON DONALDSON, AND SONG SUN 1. Introduction This is the first of a series of three articles which provide proofs of results an- nounced in [9]. Let X be a Fano manifold of complex dimension n.Letλ>0bean integer and D be a smooth divisor in the linear system |−λKX |.Forβ ∈ (0, 1) there is now a well-established notion of a K¨ahler-Einstein metric with a cone singularity of cone angle 2πβ along D. (It is often called an “edge-cone” singularity). For brevity, we will just say that ω has cone angle 2πβ along D. The Ricci curvature of such a metric ω is (1−λ(1−β))ω. Our primary concern is the case of positive Ricci −1 curvature, so we suppose throughout most of the article that β ≥ β0 > 1 − λ . However our arguments also apply to the case of non-positive Ricci curvature: see Remark 3.10. Theorem 1.1. If ω is a K¨ahler-Einstein metric with cone angle 2πβ along D with β ≥ β0,then(X, ω) is the Gromov-Hausdorff limit of a sequence of smooth K¨ahler metrics with positive Ricci curvature and with diameter bounded by a fixed number depending only on β0,λ. This suffices for most of our applications but we also prove a sharper statement. -
Semi-Supervised Deep Learning Using Improved Unsupervised Discriminant Projection?
Semi-Supervised Deep Learning Using Improved Unsupervised Discriminant Projection? Xiao Han∗, Zihao Wang∗, Enmei Tu, Gunnam Suryanarayana, and Jie Yang School of Electronics, Information and Electrical Engineering, Shanghai Jiao Tong University, China [email protected], [email protected], [email protected] Abstract. Deep learning demands a huge amount of well-labeled data to train the network parameters. How to use the least amount of labeled data to obtain the desired classification accuracy is of great practical significance, because for many real-world applications (such as medical diagnosis), it is difficult to obtain so many labeled samples. In this paper, modify the unsupervised discriminant projection algorithm from dimen- sion reduction and apply it as a regularization term to propose a new semi-supervised deep learning algorithm, which is able to utilize both the local and nonlocal distribution of abundant unlabeled samples to improve classification performance. Experiments show that given dozens of labeled samples, the proposed algorithm can train a deep network to attain satisfactory classification results. Keywords: Manifold Regularization · Semi-supervised learning · Deep learning. 1 Introduction In reality, one of the main difficulties faced by many machine learning tasks is manually tagging large amounts of data. This is especially prominent for deep learning, which usually demands a huge number of well-labeled samples. There- fore, how to use the least amount of labeled data to train a deep network has become an important topic in the area. To overcome this problem, researchers proposed that the use of a large number of unlabeled data can extract the topol- ogy of the overall data's distribution. -
Hyper-Parameter Optimization for Manifold Regularization Learning
Cassiano Ot´avio Becker Hyper-parameter Optimization for Manifold Regularization Learning Otimiza¸cao~ de Hiperparametros^ para Aprendizado do Computador por Regulariza¸cao~ em Variedades Campinas 2013 i ii Universidade Estadual de Campinas Faculdade de Engenharia Eletrica´ e de Computa¸cao~ Cassiano Ot´avio Becker Hyper-parameter Optimization for Manifold Regularization Learning Otimiza¸cao~ de Hiperparametros^ para Aprendizado do Computador por Regulariza¸cao~ em Variedades Master dissertation presented to the School of Electrical and Computer Engineering in partial fulfillment of the requirements for the degree of Master in Electrical Engineering. Concentration area: Automation. Disserta¸c~ao de Mestrado apresentada ao Programa de P´os- Gradua¸c~ao em Engenharia El´etrica da Faculdade de Engenharia El´etrica e de Computa¸c~ao da Universidade Estadual de Campinas como parte dos requisitos exigidos para a obten¸c~ao do t´ıtulo de Mestre em Engenharia El´etrica. Area´ de concentra¸c~ao: Automa¸c~ao. Orientador (Advisor): Prof. Dr. Paulo Augusto Valente Ferreira Este exemplar corresponde `a vers~ao final da disserta¸c~aodefendida pelo aluno Cassiano Ot´avio Becker, e orientada pelo Prof. Dr. Paulo Augusto Valente Ferreira. Campinas 2013 iii Ficha catalográfica Universidade Estadual de Campinas Biblioteca da Área de Engenharia e Arquitetura Rose Meire da Silva - CRB 8/5974 Becker, Cassiano Otávio, 1977- B388h BecHyper-parameter optimization for manifold regularization learning / Cassiano Otávio Becker. – Campinas, SP : [s.n.], 2013. BecOrientador: Paulo Augusto Valente Ferreira. BecDissertação (mestrado) – Universidade Estadual de Campinas, Faculdade de Engenharia Elétrica e de Computação. Bec1. Aprendizado do computador. 2. Aprendizado semi-supervisionado. 3. Otimização matemática. 4. -
2019 AMS Prize Announcements
FROM THE AMS SECRETARY 2019 Leroy P. Steele Prizes The 2019 Leroy P. Steele Prizes were presented at the 125th Annual Meeting of the AMS in Baltimore, Maryland, in January 2019. The Steele Prizes were awarded to HARUZO HIDA for Seminal Contribution to Research, to PHILIppE FLAJOLET and ROBERT SEDGEWICK for Mathematical Exposition, and to JEFF CHEEGER for Lifetime Achievement. Haruzo Hida Philippe Flajolet Robert Sedgewick Jeff Cheeger Citation for Seminal Contribution to Research: Hamadera (presently, Sakai West-ward), Japan, he received Haruzo Hida an MA (1977) and Doctor of Science (1980) from Kyoto The 2019 Leroy P. Steele Prize for Seminal Contribution to University. He did not have a thesis advisor. He held po- Research is awarded to Haruzo Hida of the University of sitions at Hokkaido University (Japan) from 1977–1987 California, Los Angeles, for his highly original paper “Ga- up to an associate professorship. He visited the Institute for Advanced Study for two years (1979–1981), though he lois representations into GL2(Zp[[X ]]) attached to ordinary cusp forms,” published in 1986 in Inventiones Mathematicae. did not have a doctoral degree in the first year there, and In this paper, Hida made the fundamental discovery the Institut des Hautes Études Scientifiques and Université that ordinary cusp forms occur in p-adic analytic families. de Paris Sud from 1984–1986. Since 1987, he has held a J.-P. Serre had observed this for Eisenstein series, but there full professorship at UCLA (and was promoted to Distin- the situation is completely explicit. The methods and per- guished Professor in 1998). -
Zero Shot Learning Via Multi-Scale Manifold Regularization
Zero Shot Learning via Multi-Scale Manifold Regularization Shay Deutsch 1,2 Soheil Kolouri 3 Kyungnam Kim 3 Yuri Owechko 3 Stefano Soatto 1 1 UCLA Vision Lab, University of California, Los Angeles, CA 90095 2 Department of Mathematics, University of California Los Angeles 3 HRL Laboratories, LLC, Malibu, CA {shaydeu@math., soatto@cs.}ucla.edu {skolouri, kkim, yowechko,}@hrl.com Abstract performed in a transductive setting, using all unlabeled data where the classification and learning process on the transfer We address zero-shot learning using a new manifold data is entirely unsupervised. alignment framework based on a localized multi-scale While our suggested approach is similar in scope to other transform on graphs. Our inference approach includes a ”visual-semantic alignment” methods (see [5, 10] and ref- smoothness criterion for a function mapping nodes on a erences therein) it is, to the best of our knowledge, the first graph (visual representation) onto a linear space (semantic to use multi-scale localized representations that respect the representation), which we optimize using multi-scale graph global (non-flat) geometry of the data space and the fine- wavelets. The robustness of the ensuing scheme allows us scale structure of the local semantic representation. More- to operate with automatically generated semantic annota- over, learning the relationship between visual features and tions, resulting in an algorithm that is entirely free of man- semantic attributes is unified into a single process, whereas ual supervision, and yet improves the state-of-the-art as in most ZSL approaches it is divided into a number of inde- measured on benchmark datasets. -
Uniqueness of Ricci Flow Solution on Non-Compact Manifolds and Integral Scalar Curvature Bound
Uniqueness of Ricci Flow Solution on Non-compact Manifolds and Integral Scalar Curvature Bound A Dissertation Presented by Xiaojie Wang to The Graduate School in Partial Fulfillment of the Requirements for the Degree of Doctor of Philosophy in Mathematics Stony Brook University August 2014 Stony Brook University The Graduate School Xiaojie Wang We, the dissertation committee for the above candidate for the Doctor of Philosophy degree, hereby recommend acceptance of this dissertation. Xiuxiong Chen – Dissertation Advisor Professor, Mathematics Department Michael Anderson – Chairperson of Defense Professor, Mathematics Department Marcus Khuri Professor, Mathematics Department Xianfeng David Gu Professor, Department of Computer Science This dissertation is accepted by the Graduate School. Charles Taber Dean of the Graduate School ii Abstract of the Dissertation Uniqueness of Ricci Flow Solution on Non-compact Manifolds and Integral Scalar Curvature Bound by Xiaojie Wang Doctor of Philosophy in Mathematics Stony Brook University 2014 In this dissertation, we prove two results. The first is about the uniqueness of Ricci flow solution. B.-L. Chen and X.-P. Zhu first proved the uniqueness of Ricci flow solution to the initial value problem by assuming bilaterally bounded curvature over the space-time. Here we show that, when the initial data has bounded curvature and is non-collapsing, the complex sectional curvature bounded from below over the space-time guarantees the short-time uniqueness of solution. The second is about the integral scalar curvature bound. A. Petrunin proved that for any complete boundary free Rieman- nian manifold, if the sectional curvature is bounded from below by negative one, then the integral of the scalar curvature over any unit ball is bounded from above by a constant depending only on the dimension. -
Semi-Supervised Deep Metric Learning Networks for Classification of Polarimetric SAR Data
remote sensing Letter Semi-Supervised Deep Metric Learning Networks for Classification of Polarimetric SAR Data Hongying Liu , Ruyi Luo, Fanhua Shang * , Xuechun Meng, Shuiping Gou and Biao Hou Key Laboratory of Intelligent Perception and Image Understanding of Ministry of Education, School of Artificial Intelligence, Xidian University, Xi’an 710071, China; [email protected] (H.L.); [email protected] (R.L.); [email protected] (X.M.); [email protected] (S.G.); [email protected] (B.H.) * Corresponding authors: [email protected] Received: 7 April 2020; Accepted: 14 May 2020; Published: 17 May 2020 Abstract: Recently, classification methods based on deep learning have attained sound results for the classification of Polarimetric synthetic aperture radar (PolSAR) data. However, they generally require a great deal of labeled data to train their models, which limits their potential real-world applications. This paper proposes a novel semi-supervised deep metric learning network (SSDMLN) for feature learning and classification of PolSAR data. Inspired by distance metric learning, we construct a network, which transforms the linear mapping of metric learning into the non-linear projection in the layer-by-layer learning. With the prior knowledge of the sample categories, the network also learns a distance metric under which all pairs of similarly labeled samples are closer and dissimilar samples have larger relative distances. Moreover, we introduce a new manifold regularization to reduce the distance between neighboring samples since they are more likely to be homogeneous. The categorizing is achieved by using a simple classifier. Several experiments on both synthetic and real-world PolSAR data from different sensors are conducted and they demonstrate the effectiveness of SSDMLN with limited labeled samples, and SSDMLN is superior to state-of-the-art methods. -
Manifold Regularization and Semi-Supervised Learning: Some Theoretical Analyses
Manifold Regularization and Semi-supervised Learning: Some Theoretical Analyses Partha Niyogi∗ January 22, 2008 Abstract Manifold regularization (Belkin, Niyogi, Sindhwani, 2004)isage- ometrically motivated framework for machine learning within which several semi-supervised algorithms have been constructed. Here we try to provide some theoretical understanding of this approach. Our main result is to expose the natural structure of a class of problems on which manifold regularization methods are helpful. We show that for such problems, no supervised learner can learn effectively. On the other hand, a manifold based learner (that knows the manifold or “learns” it from unlabeled examples) can learn with relatively few labeled examples. Our analysis follows a minimax style with an em- phasis on finite sample results (in terms of n: the number of labeled examples). These results allow us to properly interpret manifold reg- ularization and its potential use in semi-supervised learning. 1 Introduction A learning problem is specified by a probability distribution p on X × Y according to which labelled examples zi = (xi, yi) pairs are drawn and presented to a learning algorithm (estimation procedure). We are interested in an understanding of the case in which X = IRD, ∗Departments of Computer Science, Statistics, The University of Chicago 1 Y ⊂ IR but pX (the marginal distribution of p on X) is supported on some submanifold M ⊂ X. In particular, we are interested in understanding how knowledge of this submanifold may potentially help a learning algorithm1. To this end, we will consider two kinds of learning algorithms: 1. Algorithms that have no knowledge of the submanifold M but learn from (xi, yi) pairs in a purely supervised way.