Identifiez-vous pour voir le profil complet de Albert
ou
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
Identifiez-vous pour voir le profil complet de Albert
ou
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
France
Identifiez-vous pour voir le profil complet de Albert
Albert peut vous mettre en relation avec plus de 10 personnes chez Google
ou
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
3 k abonnés
+ de 500 relations
Identifiez-vous pour voir le profil complet de Albert
ou
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
Voir les relations en commun avec Albert
Albert peut vous mettre en relation avec plus de 10 personnes chez Google
ou
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
Voir les relations en commun avec Albert
ou
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
Identifiez-vous pour voir le profil complet de Albert
ou
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
À propos
Bon retour parmi nous
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
Articles de Albert
Activité
3 k abonnés
-
Albert Cohen a partagé ceciPositions in AI Systems at Google DeepMind in Paris
-
Albert Cohen a republié ceciInstitution of Engineering and Technology (IET)
Institution of Engineering and Technology (IET)
8 moisAlbert Cohen a republié ceciLast chance for this year! REACH 2025 starts on Monday, and we’re excited to be welcoming over 100 computer architecture experts to London to network, exchange ideas, and learn about the very latest research and strategies shaping computer science. We’ve got a packed programme of speakers, including: - Richard Grisenthwaite, Executive Vice President and Chief Architect, Arm, UK. - Sophie Wilson, Fellow, Broadcom, UK. - Albert Cohen, Research Scientist, Google DeepMind, France. - Michaela Blott, Senior Fellow, AMD, Ireland. - Karu Sankaralingam, Principal Research Scientist, NVIDIA, USA. - Onur Mutlu, Professor of Computer Science, ETH Zürich, Switzerland. - Partha Maji, Senior Director – AI Hardware Acceleration, Microsoft, UK. - Chris Wilkerson, Principal Engineer, Intel Corporation, USA. There’s still space for you, so don’t miss out, and make sure you have booked your seat in the room at https://spkl.io/6049AwmFh. We hope to meet you on Monday! #ComputerArchitecture #REACH2025 -
Albert Cohen a publié ceciJe participe à l’événement “REACH 2025: Reach Emerging Architectures in Computing Horizons”. Rejoignez-moi le 10 novembre.
-
Albert Cohen a republié ceciAlbert Cohen a republié ceciWe're only a few weeks away from REACH 2025 in London! The programme on Day 1 of the international computer architecture conference is packed full of highlights: ▪️Paul Kelly (Professor of Software Technology at Imperial College London) will be chairing Day 1 ▪️Richard Grisenthwaite (Executive Vice President and Chief Architect at Arm) is providing the opening keynote discussing "Arm in the Age of AI" ▪️Onur Mutlu (Professor of Computer Science at ETH Zürich) will be looking at "Memory-Centric Computing: Enabling Fundamentally-Efficient Computers" ▪️Atiq Bajwa (Chief Technology Officer & Chief Architect at Ampere Computing) provides a case study from the California-based semiconductor company ▪️Lieven Eeckhout (Full Professor at Ghent University) looks into "Sustainable Computer System Design" ▪️Partha Maji (Senior Director – AI Hardware Acceleration at Microsoft) is focusing on "Rethinking Precision: The Design Space of Block Floating-Point Formats for the LLM Era" ▪️Miquel Moretó Planas (HPC Architecture Research Area Director at Barcelona Supercomputing Center explains how they are "Designing HPC Architectures at BSC" ▪️Bill McColl (Director, Computing Systems Lab at Huawei Research Center, Zurich) looks into "Agent Machines" ▪️Boris Grot (Professor, School of Informatics at The University of Edinburgh) moderates the future-focused "Computer Architecture in 2050" panel discussion ▪️Albert Cohen (Research Scientist at Google DeepMind) leads a talk entitled "Condenser: Noun, an Apparatus for Compiling Science to the Cloud" ▪️Michaela Blott (Senior Fellow at AMD) provides the closing keynote on Day 1 focused on "Enabling the AI Revolution" See the full programme: https://lnkd.in/eNUpFvww REACH 2025 (Reach Emerging Architectures in Computing Horizons 2025) 📅10-11 November 2025 📍IET London: Savoy Place; London, UK 🔎Find out more: https://reach.theiet.org 🗣️Speakers: https://lnkd.in/e98qm_cj 👋Contact us: reach@theiet.org 🎟️Register to attend: https://lnkd.in/ecbFMmwq Institution of Engineering and Technology (IET) | IET Events and Courses | IET Venues (IET London: Savoy Place and IET Birmingham: Austin Court) #REACH2025 #ComputerArchitecture #Microprocessors #Computing #HighPerformanceComputing #LLMs #Semiconductors #AI
-
Albert Cohen a republié ceciI think formal verification is going to be used for AI evals more and more. It's traditional use case in Hardware verification is also going strong. Pretty strong set of panelists. . .Albert Cohen a republié ceciAI has many potential advantages and challenges when applied to hardware verification, especially formal verification. Join me for a free to attend webinar where we will discuss AI and hardware verification with panelists Adam Chlipala (MIT professor, formal hardware verification expert), Kanad Basu (leading work on using AI for generating formal properties) and Sean Safarpour (Executive Director at Synopsys, overseeing formal verification tools). Registration link: https://lnkd.in/gQxkUCyn
-
Albert Cohen a republié ceciAlbert Cohen a republié ceciPlus que quelques places disponibles pour la conférence Intelligence Day qui aura lieu Samedi 18 octobre à Paris ! Albert Cohen, chercheur à Google DeepMind depuis 2018, sera parmi nos intervenants experts. Les questions qu'il abordera : "L'IA générative à quel prix ? De quelle(s) performance(s) parle-t-on ? Pour quelle rationalité écologique et économique ? Pour quelle souveraineté ?" Albert Cohen dirige une équipe à la pointe de l’accélération et de l’efficacité énergétique des modèles d’intelligence artificielle. Ancien élève de l’École Normale Supérieure de Lyon et de l’Université de Versailles (Paris Saclay), agrégé de mathématiques, il rejoint l’INRIA en 2000 ainsi que l’École Polytechnique comme chargé de cours. Il a également effectué un séjour post-doctoral à l’Université de l’Illinois, puis bénéficié d’une bourse européenne Marie Curie en faveur du transfert technologique pour rejoindre Philips Research. En 2017 il rejoint le laboratoire Facebook Artificial Intelligence Research comme professeur invité. Albert est spécialiste des langages de programmation, de la compilation pour ces langages, du calcul à hautes performances et embarqué. Certains résultats ont fait l’objet de transfert technologique dans le domaine du calcul à hautes performances et de l'intelligence artificielle. 🎯 Réservez votre place : https://lnkd.in/epfcR8DZ . 📆 RDV Samedi 18 Octobre 9h-13h. Amphithéâtre de Louis-Le-Grand - 123 Rue Saint-Jacques 75005 Paris ✨ Nos autres intervenants : Cédric Vasseur 💡, Serge TISSERON , Fanchon Mayaudon Courtel , Céline TRAN , Franck Varenne , Flavien Chervet , Charbel-Raphaël Segerie .
-
Albert Cohen a partagé ceciL’intelligence artificielle est une révolution scientifique et technologique... mais aussi une transformation économique, sociale et même intime ! Le 18 octobre au matin, j’aurai le plaisir d’intervenir à l’Intelligence Day – Spécial IA, organisé par l'association Mensa Île-de-France, dans l’amphithéâtre Louis-le-Grand à Paris. Ce ne sera pas une matinée classique sur la tech. Ce sera une série d’interventions brèves, percutantes, venues de différents champs : psychologie, philosophie, informatique, robotique, cognition, inclusion… Chacun viendra partager une conviction, un doute, une vision. Pas pour convaincre, mais pour déplacer les lignes. J’interviendrai sur le thème suivant : "IA, puissance de calcul et durabilité". Samedi 18 octobre, de 9h à 13h Amphithéâtre Louis-le-Grand, 123 rue Saint-Jacques, Paris 5e Inscriptions ici : https://lnkd.in/epfcR8DZ Attention : les places sont limitées. Venez nombreux !
-
Albert Cohen a republié ceciAlbert Cohen a republié ceciMensa Île-de-France organise son prochain "Intelligence Day" le samedi 18 octobre au matin dans l'amphithéâtre du lycée Louis-Le-Grand à Paris : Comment l’IA nous impacte individuellement et collectivement Programme : - Influences Artificielles Cédric Vasseur, Mensan, conférencier IA et robotique - Éduquer à l'empathie à l’ère numérique Pr Serge Tisseron, Psychiatre, responsable du DU de Cyberpsychologie (sur les relations homme-machines) - Peut-on tomber amoureux d’un robot ? Céline Tran, Mensane, coach et coordinatrice d'intimité, auteure et conférencière - Qu'est-ce qui rend l'IA si peu explicable ? Pr Franck Varenne, philosophe des sciences - Enjeux éthiques et techniques Charbel-Raphaël Ségerie, directeur du Centre pour la sécurité de l'IA - Désinformation et cognition Julie Martinez, essayiste, avocate de formation, DG de France Positive et spécialiste des fake-news - Les arbres ne montent jamais jusqu'au ciel, les systèmes d'IA sont faits du même bois Albert Cohen, chercheur chez Google DeepMind - IA et créativité humaine Flavien Chervet, Conférencier et YouTubeur - IA, inclusion et identitésFanchon Mayaudon Courtel, experte IA & défenseuse des droits LGBT+ et un invité surprise : Leenby Le tarif public est de 8 euros https://lnkd.in/epfcR8DZ A bientôt !
-
Albert Cohen a partagé ceciI'll be attending the amazing HPCA/PPoPP/CGO 2025. Let me know if you're planning to attend so that we can say Hi! 👋 Or register now and join me at the event! https://lnkd.in/eTZSti5D #hpca2025 #cgo2025 #ppopp2025 #cc2025 - via #Whova event appHPCA/PPoPP/CGO/CC/NVMW 2025 Registration hosted by WhovaHPCA/PPoPP/CGO/CC/NVMW 2025 Registration hosted by Whova
-
Albert Cohen a aimé ceciAlbert Cohen a aimé ceciHonorée d'être intervenue dans le Conseil Présidentiel de la Science à la Présidence de la République, grâce à la suggestion d'Anne-Marie Kermarrec, à côté de David Chavalarias, Pierre Paul Zalio, Hugo Duminil-Copin et plusieurs collègues, pour discuter notamment de la #désinformation, du #factchecking, de l'#IA comme vecteur d'attaque et de défense dans l'espace informationnel. L'occasion d'évoquer notre collaboration avec Radio France sur StatCheck, outil de vérification d'affirmation statistiques, ainsi que le crawler intelligent qui permet de retrouver des statistiques publiées dans les méandres d'un site Web. Recherches menées avec notamment Oana Balalau, Antoine GAUQUIER, Helena Galhardas, Pierre Senellart Merci et bravo aussi à Oana Goga, Asmaa EL FRAIHI, Olivier Blazy et ses collègues pour les idées et suggestions! ✨ 🚀
-
Albert Cohen a aimé ceciAlbert Cohen a aimé ceciWe just finished the 2026 edition of our VLSI course (https://vlsi.ethz.ch) that resulted in 37 ASICs, eight of which we could afford to send out for fabrication in IHP130nm 🎉 🥂 😀Teaching practical IC Design using Open-Source EDATeaching practical IC Design using Open-Source EDAFrank Kagan Gurkaynak
-
Albert Cohen a aimé ceciAlbert Cohen a aimé ceciKalray annonce son chiffre d'affaires du premier semestre 2026 : 🔹Croissance de +58% de l'activité "semi-conducteur" 🔹Chiffre d'affaires consolidé de 8,5 M€ 🔹EBITDA attendu largement positif 🔹Signature du contrat majeur de collaboration stratégique avec un leader des infrastructures IA et HPC 🔹Confirmation des perspectives financières 2026 : croissance à deux chiffres du chiffre d’affaires en nouvelle amélioration de l’EBITDA. ➡️ Le premier semestre 2026 marque une étape majeure dans la transformation de Kalray. En 18 mois, Kalray a profondément transformé son modèle économique et démontre aujourd'hui sa capacité à conjuguer croissance, amélioration de sa rentabilité et création de valeur à partir de ses technologies et de son expertise, dans un monde révolutionné par l’IA. ➡️ La signature de son partenariat avec un leader des infrastructures IA et HPC, ainsi que l'extension de sa collaboration avec Openchip, illustrent la capacité de Kalray à convaincre des acteurs de premier plan d'intégrer ses technologies au cœur de leurs prochaines générations de puces et de bâtir des relations sur le long terme. Ces accords valident la pertinence du positionnement de Kalray et renforcent sa visibilité pour les années à venir. Dans un contexte où la souveraineté technologique européenne devient un enjeu majeur, Kalray entend jouer un rôle de partenaire de référence pour les infrastructures de calcul intensif et d'intelligence artificielle.
-
Albert Cohen a aimé ceciNew openings in our team with Albert Cohen and Ulysse Beaugnon! If you are interested in high-impact projects where machine learning meets formal methods, compiler construction and hardware design, all of this with exceptionally skilled and kind colleagues, please consider applying.
-
Albert Cohen a aimé ceciAlbert Cohen a aimé ceciTPU opcodes are not public but the compiler's model of the machine is. NVIDIA runs the other way. The instruction set is documented, and the ptxas compiler is not. And NVIDIA's documentation stops higher up than people assume: PTX is a virtual ISA, but SASS is published simply as a list of opcode names. So I went through the open compiler sources, OpenXLA and the Mosaic dialect inside JAX. They name the parts of the machine they generates code for, and there is useful information in there for anyone writing TPU kernels. tpu.log is printf. A tag string plus whatever values you hand it. Next to it sit log_buffer, and trace_start/trace_stop for timing regions. I feel kernel debugging on TPU has a worse reputation than it deserves. tpu.weird takes an f32 and returns a bool. JAX's lowering evaluates lax.is_finite is its negation. So the TPU has a hardware predicate for "this float is weird." And the MXU is a FIFO. Push the weights, stream the activations, pop the result. Weight-stationary, more than one per core, 32-bit accumulator enforced by the verifier. "bf16 in, fp32 accumulate" is the only shape the unit offers. 90 operations, eight memory spaces, and a lot of guidance to developers. https://lnkd.in/eCU3xkF4
-
Albert Cohen a aimé ceciAlbert Cohen a aimé ceciEurope has a serious problem of sovereignty, but nobody seems to careEurope has a serious problem of sovereignty, but nobody seems to careSandro D'Elia
-
Albert Cohen a aimé ceciAlbert Cohen a aimé ceciGreat to be back at DAC in Long Beach, meeting many colleagues and partners, and catching up with the latest AI technologies in EDA. Big congrats to ICE's Chiara Ghinami with her successful presentation on Embedded Fuzzing using SystemC Virtual Prototypes. MachineWare GmbH
-
Albert Cohen a aimé ceciToday, I feel immensely proud, both as an alumnus of École Polytechnique and as its Provost, to see Hong Wang, alumnus and doctor Honoris Causa of Ecole polytechnique receive the prestigious Fields Medal. This outstanding achievement once again shines a spotlight on mathematics and celebrates fundamental research. It also reflects Ecole Polytechnique deep commitment to advancing mathematical sciences at the highest level. My warmest congratulations to Hong Wang on this remarkable recognition. Her accomplishment will undoubtedly inspire future generations of mathematicians and researchers.Albert Cohen a aimé ceci🇫🇷 C'est une première historique : l'une de nos alumni a reçu la médaille Fields 2026 à l'occasion du Congrès international des mathématiciens. 🎉 Toutes nos félicitations à Hong Wang (X2010 et Doctor Honoris Causa de l'X), lauréate de la médaille Fields, la plus haute distinction internationale en mathématiques ! Première diplômée de l'École polytechnique à recevoir cette récompense d'exception, Hong Wang incarne l'excellence scientifique ainsi que les valeurs qui animent la recherche : la curiosité, la rigueur, la persévérance et l'audace intellectuelle, indispensables aux plus grandes avancées. Après le Clay Research Award et le New Horizons in Mathematics Prize, cette médaille Fields vient couronner ses travaux remarquables en analyse harmonique, dont les contributions repoussent les frontières de la connaissance 🤩 Hong Wang, encore toutes nos félicitations pour cette réussite exceptionnelle qui entre dans l'histoire et constitue une formidable source d'inspiration pour les générations futures ! 👏 Retrouvez l'actu ➡️ https://lnkd.in/echP533h Zoom sur la conjecture de Kakeya ➡️ https://lnkd.in/eFCt35JD --- 🇬🇧 A first in our history: one of our alumni has been awarded the Fields Medal...! Congratulations to Hong Wang (X2010 and Honorary Doctor of École Polytechnique) on being awarded the 2026 Fields Medal, the highest international distinction in mathematics! 🎉 As the very first École Polytechnique's alumna to receive this exceptional honor, Hong Wang embodies scientific excellence and the values that drive groundbreaking research: curiosity, rigor, perseverance, and intellectual boldness. Following the Clay Research Award and the New Horizons in Mathematics Prize, the Fields Medal recognizes her outstanding contributions to harmonic analysis, advancing the frontiers of mathematical knowledge 🤩 Hong Wang, congratulations once again on this extraordinary achievement. Your success marks a historic milestone and will inspire generations of scientists to come. 👏 A closer look at Kakeya's conjecture ➡️ https://lnkd.in/enScpu5F #EducationLX I #FormationLX I Institut des Hautes Études Scientifiques - IHES I Institut Polytechnique de Paris I Polytechnique Alumni
Expérience et formation
-
Google
******** *********
-
*****
******** *********
-
********
******** *********
-
***** ******* ********** ** ****
****** ** ******* * ** ******** ******* undefined
-
-
********** ** ********** *************************
*** ******** *******
-
Voir toute l’expérience de Albert
Découvrez son poste, son ancienneté et plus encore.
Bon retour parmi nous
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
Nouveau sur LinkedIn ? Inscrivez-vous maintenant
ou
En cliquant sur Continuer pour vous inscrire ou vous identifier, vous acceptez les Conditions d’utilisation, la Politique de confidentialité et la Politique relative aux cookies de LinkedIn.
Projets
Langues
-
French
-
-
English
-
Organisations
-
ACM
-
Recommandations reçues
3 personnes ont recommandé Albert
Inscrivez-vous pour y accéderVoir le profil complet de Albert
-
Découvrir vos relations en commun
-
Être mis en relation
-
Contacter Albert directement
Autres profils similaires
-
Laércio Lima Pilla
Laércio Lima Pilla
Laboratoire Bordelais de Recherche en Informatique (LaBRI)
582 abonnésTalence -
Jean-François Mascari
Jean-François Mascari
Mathématiques et Interactions - Université Côte d'Azur
3 k abonnésNice
Découvrir plus de posts
-
Stephen Pimentel
Independent • 4 k abonnés
A team of Claude Code agents with Opus 4.6 can execute large, complex software projects without human supervision when supported by harnesses and tests. Sixteen parallel agents ran continuously in isolated containers, coordinated through a simple locking mechanism on a shared repository, and collectively produced a clean room Rust C compiler capable of building the Linux 6.9 kernel across multiple architectures. The key insight was that sustained autonomy depends less on prompting and more on environmental design, especially high quality tests, concise and machine readable feedback, and structures that make parallel progress possible. Deterministic subsampled testing, aggressive regression prevention, and explicit project state documentation allowed agents to orient themselves and recover from errors. Parallelism enabled specialization, but only when large tasks were decomposed using external oracles to avoid duplication of effort. The experiment demonstrated that current models can reach the edge of end to end systems programming, while also exposing limits in robustness, efficiency, and architectural judgment that frequent changes still destabilize. https://lnkd.in/gWNvhx7D
6
-
HGPU group
HGPU group • 349 abonnés
Inside VOLT: Designing an Open-Source GPU Compiler (Tool) Recent efforts in open-source GPU research are opening new avenues in a domain that has long been tightly coupled with a few commercial vendors. Emerging open GPU architectures define SIMT functionality through their own ISAs, but executing existing GPU programs and optimizing performance on these ISAs relies on a compiler framework that is technically complex and often undercounted in open-hardware development costs....
20
-
Leandro Aolita
Technology Innovation… • 1 k abonnés
Impressive milestone of the Technology Innovation Institute's quantum-inspired algorithms and quantum middleware teams together with NVIDIA. This is what happens when strong algorithms theory (graph tensor networks plus belief propagation) meets power GPU-acceleration. Thanks in particular to Ilia Luchnikov, Daniel Vieira, Stefano Carrazza, Stefano Mensa, PhD MBCS, Esperanza Cuenca Gómez, Mazen Khalil, Nicholas Harrigan, and Elica (Elitsa) Kyoseva. See https://lnkd.in/daynMAZh for details.
56
2 commentaires -
Nickey Chen (She/Her/Hers)
573 abonnés
The PyTorch 2.9 release delivers measurable performance gains on Arm CPUs, with contributions from Arm’s engineering teams across core areas of the stack. This includes optimizations through oneDNN and OpenBLAS, improved operator coverage, and stronger compiler consistency, all designed to deliver faster, more stable AI workloads on Arm platforms. These updates are part of Arm’s ongoing collaboration with the PyTorch community to enable open, efficient, and scalable AI performance for developers everywhere. Read more about what’s new and see you next week at PyTorch Conference: https://okt.to/91w4kU
8
-
Compilers Lab
9 k abonnés
In this new video (https://lnkd.in/dumY28mt), Fernando Pereira presents ongoing work at UFMG's Compilers Lab focused on the design and implementation of more efficient tensor compilers. A tensor compiler is a tool that translates machine learning models, expressed in high-level representations such as Torch, TensorFlow, or ONNX, into low-level, highly optimized machine code. Over the past two years, researchers at UFMG's Compilers Lab, in cooperation with Universidade Estadual de Campinas and Cadence have been developing improved kernel fusion algorithms and autotuning techniques. Kernel fusion is a classic compiler optimization that combines two or more sequences of loops into a single loop. In modern tensor compilers, this optimization can also be viewed as a variation of instruction selection, where the implementations of multiple kernels are replaced by a single kernel that preserves their combined semantics. We have designed and implemented a new kernel fusion algorithm that currently runs in the XNNC compiler from Cadence Design Systems [1]. Autotuning is a profile-guided optimization technique in which the compiler runs a program, collects profile data, and uses that information to select the best optimizations for the program before recompiling it. This feedback loop continues until performance converges to an acceptable level. We have designed an algorithm, called Droplet Search, that maps the search space of optimization parameters (such as unrolling factor, tiling window, and number of threads) to a coordinate space. Droplet Search then applies Coordinate Descent to find the best combination of optimization parameters for a given kernel [2, 3]. This algorithm is currently integrated into Apache TVM. References: [1] Michael Canesche, Vanderson Rosario, Edson Borin, Fernando Pereira: Fusion of Operators of Computational Graphs via Greedy Clustering: The XNNC Experience. CC 2025: 117-127 (https://lnkd.in/dsQJtSkz) [2] Michael Canesche, Vanderson Martins do Rosario, Edson Borin, Fernando Magno Quintao Pereira: The Droplet Search Algorithm for Kernel Scheduling. ACM Trans. Archit. Code Optim. 21(2): 35, 2024 (https://lnkd.in/dBfQid9z). [3] Michael Canesche, Gaurav Verma, Fernando Magno Quintao Pereira: Explore as a Storm, Exploit as a Raindrop: On the Benefit of Fine-Tuning Kernel Schedulers with Coordinate Descent. CoRR abs/2406.20037, 2024 (https://lnkd.in/dgxxfdYR). #compiler #research #academia #tensorcompiler #machinelearning
79
-
Jeremy Fowers
AMD • 1 k abonnés
🚀 500 → 645 stars on 🍋 Lemonade repo in the time it took to write this post! Developers worldwide are sending a clear message: local LLMs need to be… 1️⃣ Optimized for each PC (NPU, ROCm, etc.) 2️⃣ Zero-friction for users and developers alike 3️⃣ Committed to open-source principles The team at AMD, alongside a fast-growing community of contributors and mentorship from Open Core Ventures Catalyst, is delivering on all three ✨ 👉 Star the repo: https://lnkd.in/ebfd_f75 👉 Join the conversation: https://lnkd.in/e8unnwXD Huge thanks to everyone making this possible, including: Victoria Godsoe Daniel Holanda Noronha Ramakrishnan Sivakumar Tomasz Iniewicz Kalin Ovtcharov Iswarya Alex Mike Kraus Somayeh Rahimipour Adrian Macias Richter Edan Sasson Joseph Melber Phil James-Roxby Ramine Roane Hisham Chowdhury Tyler Straub Kevin Ding Matthew Letts Joe Pizzini Ryan Radjabi Hyunho Ahn Patrick Worfolk Gabriel Weisz Rajeev P. Ashish Sirasao Alex Smith #LocalLLM #OpenSourceAI #technology #developers #AMD
47
9 commentaires -
Ashwath Aithal
NVIDIA • 7 k abonnés
NeMo RL now supports the Decoupled Clip and Dynamic Sampling Policy Optimization (DAPO) algorithm. DAPO enhances the Generalized Reinforcement Policy Optimization (GRPO) by incorporating several advanced features: - Clip-Higher - Dynamic Sampling - Token-Level Policy Gradient Loss - Overlong Reward Shaping These additions aim to provide more stable and efficient reinforcement learning training. For further details, refer to the DAPO guide. https://lnkd.in/geGgziXy
112
-
Priyanshu Kumar
Birla Institute of… • 1 k abonnés
Most LLM evaluation pipelines rely on static test sets or uniform sampling, allocating equal effort across all prompt categories. This leads to inefficient use of evaluation budget and slower failure discovery. I modeled evaluation as a resource allocation problem and applied multi-armed bandit algorithms (UCB1 and decaying epsilon-greedy) to dynamically prioritize high-risk prompt regions. Across 30 independent runs (20,000 plays each): UCB1 final regret: ~104 Epsilon-Greedy (decaying): ~12 p = 0.0012 (paired t-test) This shows a statistically significant advantage for adaptive exploration in concentrating evaluation effort where failures are more likely. https://lnkd.in/g6jDS4M4 #MachineLearning #LLM #Algorithms
13
5 commentaires -
Furu Wei
Microsoft Research Asia • 13 k abonnés
Introducing Generative Adversarial Distillation (GAD): a novel GAN-style formulation and framework that facilitates both on-policy and black-box distillation of large language models (LLMs). GAD is the first technique to enable block-box on-policy distillation from proprietary teachers where internal logits or parameters are inaccessible, or distillation between teacher and student LLMs with incompatible vocabularies. GAD expands our prior work on white-box on-policy distillation (i.e., MiniLLM), pioneering block-box on-policy distillation for LLM training. Specifically, GAD frames the student LLM as a generator and trains a discriminator to distinguish its responses from the teacher LLM’s, creating a minimax game. The discriminator acts as an on-policy reward model that co-evolves with the student, providing stable, adaptive feedback. Experimental results show that GAD consistently surpasses the commonly used sequence-level knowledge distillation. In particular, Qwen2.5-14B-Instruct (student) trained with GAD becomes comparable to its teacher, GPT-5-Chat, on the LMSYS-Chat automatic evaluation. The results establish GAD as a promising and effective paradigm for black-box LLM distillation. Our team has been conducting fundamental research in knowledge distillation with wide adoptions across the industry. - MiniLM: We introduced multi-head attention distillation, establishing the most effective distillation method for BERT-style models. The open-source MiniLM models (e.g., 6x384) have become the most widely utilized small encoder models on the Hugging Face. - MiniLLM: Our proposed Reverse KLD is recognized as one of the most effective, de facto on-policy distillation approaches for modern LLM training, which has been widely used by Thinking Machines, Gemma, and many other teams and models. - BitDistill: We proposed BitNet Distillation to finetune off-the-shelf full-precision LLMs (e.g., Qwen) into 1.58-bit precision (ternary weights {-1, 0, 1}), achieving performance parity with the full-precision counterparts on specific downstream tasks. - GAD: The development of Generative Adversarial Distillation (GAD) now allows for black-box on-policy distillation, overcoming two major prior limitations: (1) Distillation from proprietary teachers where internal logits or parameters are inaccessible; (2) Distillation between teacher and student LLMs with incompatible vocabularies. https://lnkd.in/gMaP2c7w
126
2 commentaires