On Learning Matrices with Orthogonal Columns or Disjoint Supports

Vervier, Kevin; Mahé, Pierre; D’Aspremont, Alexandre; Veyrieras, Jean-Baptiste; Vert, Jean-Philippe

doi:10.1007/978-3-662-44845-8_18

Kevin Vervier^23,24,25,
Pierre Mahé²³,
Alexandre D’Aspremont²⁶,
Jean-Baptiste Veyrieras²³ &
…
Jean-Philippe Vert^24,25

Part of the book series: Lecture Notes in Computer Science ((LNAI,volume 8726))

Included in the following conference series:

Joint European Conference on Machine Learning and Knowledge Discovery in Databases

2869 Accesses

Abstract

We investigate new matrix penalties to jointly learn linear models with orthogonality constraints, generalizing the work of Xiao et al. [24] who proposed a strictly convex matrix norm for orthogonal transfer. We show that this norm converges to a particular atomic norm when its convexity parameter decreases, leading to new algorithmic solutions to minimize it. We also investigate concave formulations of this norm, corresponding to more aggressive strategies to induce orthogonality, and show how these penalties can also be used to learn sparse models with disjoint supports.

Download to read the full chapter text

Chapter PDF

Matrix completion with nonconvex regularization: spectral operators and scalable algorithms

Article 14 March 2020

Online optimization for max-norm regularization

Article 07 February 2017

An Overview of Computational Sparse Models and Their Applications in Artificial Intelligence

Keywords

These keywords were added by machine and not by the authors. This process is experimental and the keywords may be updated as the learning algorithm improves.

References

Bach, F., Jenatton, R., Mairal, J., Obozinski, G.: Optimization with sparsity-inducing penalties. Foundations and Trends® in Machine Learning 4(1), 1–106 (2011)
Article MATH Google Scholar
Bakker, B., Heskes, T.: Task clustering and gating for bayesian multitask learning. J. Mach. Learn. Res. 4, 83–99 (2003)
Google Scholar
Barvinok, A. A Course in Convexity. American Mathematical Society (2002)
Google Scholar
Baxter, J.: A model of inductive bias learning. Journal of Artificial Intelligence Research 12, 149–198 (2000)
MATH MathSciNet Google Scholar
Borwein, J.M., Lewis, A.S.: Convex Analysis and Nonlinear Optimization: Theory and Examples. Cms Books in Mathematics Series. Springer (2000)
Google Scholar
Boyd, S., Vandenberghe, L.: Convex Optimization. Cambridge University Press (2004)
Google Scholar
Brickman, L.: On the field of values of a matrix. Proceedings of the American Mathematical Society, 61–66 (1961)
Google Scholar
Cai, L., Hofmann, T.: Hierarchical document categorization with support vector machines. In: Proceedings of the Thirteenth ACM International Conference on Information and Knowledge Management, pp. 78–87. ACM, New York (2004)
Google Scholar
Calder, A.J., Burton, A.M., Miller, P., Young, A.W., Akamatsu, S.: A principal component analysis of facial expressions. Vision Res. 41(9), 1179–1208 (2001)
Article Google Scholar
Caruana, R.: Multitask learning. Mach. Learn. 28(1), 41–75 (1997)
Article MathSciNet Google Scholar
Chandrasekaran, V., Recht, B., Parrilo, P.A., Willsky, A.S.: The convex geometry of linear inverse problems. Found. Comput. Math. 12(6), 805–849 (2012)
Article MATH MathSciNet Google Scholar
Evgeniou, T., Micchelli, C., Pontil, M.: Learning multiple tasks with kernel methods. J. Mach. Learn. Res. 6, 615–637 (2005)
MATH MathSciNet Google Scholar
Hwang, S.J.J., Grauman, K., Sha, F.: Learning a tree of metrics with disjoint visual features. In: Shawe-Taylor, J., Zemel, R.S., Bartlett, P., Pereira, F.C.N., Weinberger, K.Q. (eds.) Adv. Neural. Inform. Process Syst. 24, pp. 621–629 (2011)
Google Scholar
Jacob, L., Vert, J.-P.: Protein-ligand interaction prediction: an improved chemogenomics approach. Bioinformatics 24(19), 2149–2156 (2008)
Article Google Scholar
Lovász, L., Schrijver, A.: Cones of matrices and set-functions and 0-1 optimization. SIAM Journal on Optimization 1(2), 166–190 (1991)
Article MATH MathSciNet Google Scholar
McCallum, A., Rosenfeld, R., Mitchell, T.M., Ng, A.Y.: Improving text classification by shrinkage in a hierarchy of classes. In: Proceedings of the Fifteenth International Conference on Machine Learning, pp. 359–367. Morgan Kaufmann Publishers Inc., San Francisco (1998)
Google Scholar
Obozinski, G., Taskar, B., Jordan, M.I.: Joint covariate selection and joint subspace selection for multiple classification problems. Statistics and Computing 20(2), 231–252 (2010)
Article MathSciNet Google Scholar
Rockafellar, R.T.: Convex Analysis. Princeton University Press, Princeton (1970)
MATH Google Scholar
Romera-Paredes, B., Argyriou, A., Berthouze, N., Pontil, M.: Exploiting unrelated tasks in multi-task learning. J. Mach. Learn. Res. - Proceedings Track 22, 951–959 (2012)
Google Scholar
Shor, N.Z.: Quadratic optimization problems. Soviet Journal of Computer and Systems Sciences 25, 1–11 (1987)
MATH MathSciNet Google Scholar
Srebro, N., Rennie, J.D.M., Jaakkola, T.S.: Maximum-margin matrix factorization. In: Saul, L.K., Weiss, Y., Bottou, L. (eds.) Adv. Neural. Inform. Process Syst. 17, pp. 1329–1336. MIT Press, Cambridge (2005)
Google Scholar
Thrun, S., Pratt, L. (eds.): Learning to learn. Kluwer Academic Publishers, Norwell (1998)
MATH Google Scholar
Tibshirani, R.: Regression shrinkage and selection via the lasso. J. R. Stat. Soc. Ser. B 58(1), 267–288 (1996)
MATH MathSciNet Google Scholar
Xiao, L., Zhou, D., Wu, M.: Hierarchical classification via orthogonal transfer. In: Getoor, L., Scheffer, T. (eds.) Proceedings of the 28th International Conference on Machine Learning, ICML 2011, Bellevue, Washington, USA, June 28-July 2, pp. 801–808. Omnipress (2011)
Google Scholar
Xiao, L.: Dual averaging methods for regularized stochastic learning and online optimization. J. Mach. Learn. Res. 9999, 2543–2596 (2010)
Google Scholar

Download references

Author information

Authors and Affiliations

Data and Knowledge Lab, Biomerieux, 69280, Marcy l’Etoile, France
Kevin Vervier, Pierre Mahé & Jean-Baptiste Veyrieras
Centre for Computational Biology, Mines ParisTech, 77300, Fontainebleau, France
Kevin Vervier & Jean-Philippe Vert
Institut Curie, INSERM U900, 75005, Paris, France
Kevin Vervier & Jean-Philippe Vert
CNRS and D.I. UMR 8548, Ecole normale supérieure, 75005, Paris, France
Alexandre D’Aspremont

Authors

Kevin Vervier
View author publications
You can also search for this author in PubMed Google Scholar
Pierre Mahé
View author publications
You can also search for this author in PubMed Google Scholar
Alexandre D’Aspremont
View author publications
You can also search for this author in PubMed Google Scholar
Jean-Baptiste Veyrieras
View author publications
You can also search for this author in PubMed Google Scholar
Jean-Philippe Vert
View author publications
You can also search for this author in PubMed Google Scholar

Editor information

Editors and Affiliations

Faculty of Applied Sciences, Department of Computer and Decision Engineering, Université Libre de Bruxelles, Av. F. Roosevelt, CP 165/15, 1050, Brussels, Belgium
Toon Calders
Dipartimento di Informatica, Università degli Studi “Aldo Moro”, via Orabona 4, 70125, Bari, Italy
Floriana Esposito
Department of Computer Science, Universität Paderborn, Warburger Str. 100, 33098, Paderborn, Germany
Eyke Hüllermeier
Dipartimento di Informatica,, Università degli Studi di Torino, Corso Svizzera 185, 10149, Torino, Italy
Rosa Meo

Rights and permissions

Reprints and permissions

Copyright information

About this paper

Cite this paper

Vervier, K., Mahé, P., D’Aspremont, A., Veyrieras, JB., Vert, JP. (2014). On Learning Matrices with Orthogonal Columns or Disjoint Supports. In: Calders, T., Esposito, F., Hüllermeier, E., Meo, R. (eds) Machine Learning and Knowledge Discovery in Databases. ECML PKDD 2014. Lecture Notes in Computer Science(), vol 8726. Springer, Berlin, Heidelberg. https://doi.org/10.1007/978-3-662-44845-8_18

Download citation

DOI: https://doi.org/10.1007/978-3-662-44845-8_18
Publisher Name: Springer, Berlin, Heidelberg
Print ISBN: 978-3-662-44844-1
Online ISBN: 978-3-662-44845-8
eBook Packages: Computer ScienceComputer Science (R0)

Publish with us

Policies and ethics

On Learning Matrices with Orthogonal Columns or Disjoint Supports

Abstract

Chapter PDF

Similar content being viewed by others

Matrix completion with nonconvex regularization: spectral operators and scalable algorithms

Online optimization for max-norm regularization

An Overview of Computational Sparse Models and Their Applications in Artificial Intelligence

Keywords

References

Author information

Authors and Affiliations

Editor information

Editors and Affiliations

Rights and permissions

Copyright information

About this paper

Cite this paper

Download citation

Publish with us

Navigation

On Learning Matrices with Orthogonal Columns or Disjoint Supports

Abstract

Chapter PDF

Similar content being viewed by others

Matrix completion with nonconvex regularization: spectral operators and scalable algorithms

Online optimization for max-norm regularization

An Overview of Computational Sparse Models and Their Applications in Artificial Intelligence

Keywords

References

Author information

Authors and Affiliations

Editor information

Editors and Affiliations

Rights and permissions

Copyright information

About this paper

Cite this paper

Download citation

Share this paper

Publish with us

Search

Navigation