Sign in

Pedro Sarmento

@umpedronosapato.bsky.social
288 followers 343 following 39 posts

AI & Music Data Scientist at @Music.AI | prev. @c4dm

PostsRepliesMedia
Pedro Sarmento @umpedronosapato.bsky.social · 11/04/2025
can't get enough of guitar-MIR 🎸
010
Pedro Sarmento @umpedronosapato.bsky.social · 03/04/2025
Really creative use of AI for video by #meatdept on this banger (non AI) release by #igorrr 🤘 www.youtube.com/watch?v=rbkk...
youtube.com
Igorrr - ADHD (Official Video)
YouTube video by Metal Blade Records
010
Pedro Sarmento @umpedronosapato.bsky.social · 28/03/2025
Let us hear your AI-assisted bangers 🤘
020
Pedro Sarmento @umpedronosapato.bsky.social · 25/03/2025
So many great works 🤘
020
Reposted by Pedro Sarmento
C4DM at QMUL @c4dm.bsky.social · 22/03/2025
An exciting novel contribution by our student @jinhua-liang.bsky.social, supervised by @emmanouilb.bsky.social
031
Pedro Sarmento @umpedronosapato.bsky.social · 21/03/2025
Good luck to all the titans submitting to #ISMIR2025 🤘excited to see what this year's edition will bring 🎸
130
Reposted by Pedro Sarmento
Nigel Warburton @nigelwarburton.bsky.social · 11/03/2025
4’33, One Minute, and the copyright grab - my Everyday Philosophy column @theneweuropean.bsky.social www.theneweuropean.co.uk/nigel-warbur...
theneweuropean.co.uk
Everyday Philosophy: John Cage and the sound of silence
A collective called the 1000 Artists have followed in the composer’s footsteps by releasing a silent protest album
0196
Pedro Sarmento @umpedronosapato.bsky.social · 10/03/2025
I'm running a paid study on guitar timbre transfer - it should take approximately 30min 🎸 If you're interested, please reach out via DM!
023
Reposted by Pedro Sarmento
Oriol (Uri) Nieto @urinieto.bsky.social · 04/03/2025
I love how DiffRhythm keeps changing time signatures à la Dream Theater (ie, seemingly random). The vocals are in a quite deep uncanny valley, but the music sounds super good. And the audio prompting works really well! And all open source! Great job, titans <3 huggingface.co/spaces/ASLP-...
huggingface.co
DiffRhythm - a Hugging Face Space by ASLP-lab
Blazingly Fast and Embarrassingly Simple Song Generation
061
Pedro Sarmento @umpedronosapato.bsky.social · 22/02/2025
They're out 🤘
030
Reposted by Pedro Sarmento
Scott H. Hawley @drscotthawley.bsky.social · 12/02/2025
Video of @stefanlattner.bsky.social 's talk at DMRN+19 is finally online: "Models of Musical Signals: Representation, Learning & Generation" @c4dm.bsky.social www.youtube.com/watch?v=ixHf...
youtube.com
Models of Musical Signals: Representation, Learning & Generation. Stefan Lattner (Sony SCL). DMRN+19
YouTube video by C4DM - Centre for Digital Music
092
Reposted by Pedro Sarmento
Sander Dieleman @sedielem.bsky.social · 10/02/2025
Great interview with @jascha.sohldickstein.com about diffusion models! This is the first in a series: similar interviews with Yang Song and yours truly will follow soon. (One of these is not like the others -- both of them basically invented the field, and I occasionally write a blog post 🥲)
youtube.com
History of Diffusion - Jascha Sohl-Dickstein
YouTube video by Bain Capital Ventures
04311
Reposted by Pedro Sarmento
kaseypocius.bsky.social @kaseypocius.bsky.social · 07/02/2025
exitpoints.bandcamp.com/album/you-ar... Grab some albums on bandcamp today, support independent artists and Musicares!
exitpoints.bandcamp.com
You Are The Right Length, by Exit Points
10 track album
261
Pedro Sarmento @umpedronosapato.bsky.social · 07/02/2025
Very excited to share our latest work, the GigaMIDI dataset with > 1.4M files, published at #TISMIR 🤘 It was a huge pleasure to collaborate with such a team of titans transactions.ismir.net/articles/10....
transactions.ismir.net
The GigaMIDI Dataset with Features for Expressive Music Performance Detection | Transactions of the International Society for Music Information Retrieval
The Transactions of the International Society for Music Information Retrieval publishes novel scientific research in the field of music information retrieval (MIR), an interdisciplinary research area concerned with processing, analysing, organising and accessing music information. We welcome submissions from a wide range of disciplines, including computer science, musicology, cognitive science, library & information science and electrical engineering.TISMIR was established to complement the widely cited ISMIR conference proceedings and provide a vehicle for the dissemination of the highest quality and most substantial scientific research in MIR. TISMIR retains the Open Access model of the ISMIR Conference proceedings, providing rapid access, free of charge, to all journal content. In order to encourage reproducibility of the published research papers, we provide facilities for archiving the software and data used in the research. To avoid excessive cost to the authors or their institutions, TISMIR is published in electronic-only format.
0123
Reposted by Pedro Sarmento
C4DM at QMUL @c4dm.bsky.social · 05/02/2025
From the 25th February to 4th March 2025, two C4DM researchers will participate at the 39th Annual AAAI Conference on Artificial Intelligence (AAAI 2025). More info at: www.c4dm.eecs.qmul.ac.uk/news/2025-02...
c4dm.eecs.qmul.ac.uk
The following works were authored/coauthored by C4DM PhD students and academic staff:
053
Pedro Sarmento @umpedronosapato.bsky.social · 03/02/2025
this is pricelessly sad and great at the same time 🤘 Courtney LaPlante is such a titan
110
Reposted by Pedro Sarmento
Yoshua Bengio @yoshuabengio.bsky.social · 29/01/2025
Today, we are publishing the first-ever International AI Safety Report, backed by 30 countries and the OECD, UN, and EU. It summarises the state of the science on AI capabilities and risks, and how to mitigate those risks. 🧵 Full Report: assets.publishing.service.gov.uk/media/679a0c... 1/21
7256104
Pedro Sarmento @umpedronosapato.bsky.social · 28/01/2025
Another banger 🤘
000
Pedro Sarmento @umpedronosapato.bsky.social · 28/01/2025
Following up on the release of open source models that are shaking the AI status quo: YuE (乐) 🎵 - full music generation - demo: map-yue.github.io - conditioned on lyrics (even does vocal fry and growls 🤘) - Non-commercial license Super impressive and disruptive work! github.com/multimodal-a...
github.com
GitHub - multimodal-art-projection/YuE: YuE: Open Full-song Generation Foundation Model, something similar to Suno.ai but open
YuE: Open Full-song Generation Foundation Model, something similar to Suno.ai but open - multimodal-art-projection/YuE
040
Reposted by Pedro Sarmento
Deezer Research @researchdeezer.bsky.social · 27/01/2025
We are proudly engaged in improving transparency both for artists and users relative to the spread of AI generated music on our platform. Based on months of research we're deploying a large scale detector and aim to remove such content from our recommendations: newsroom-deezer.com/2025/01/deez...
newsroom-deezer.com
Deezer deploys cutting-edge AI detection tool for music streaming - Deezer Newsroom
Paris, January 24, 2025 – Deezer (Paris Euronext: DEEZR), the global music experiences platform has deployed a cutting-edge AI music detection tool, discovering that roughly 10,000 fully AI generated ...
0288
Reposted by Pedro Sarmento
AES AIMLA 2025 @aesaimla25.bsky.social · 27/01/2025
📢 Call for contributions: First AES International Conference on Artificial Intelligence and Machine Learning for Audio (AIMLA 2025), London, Sept. 8-10, 2025. More info: aes2.org/contribution... @c4dm.bsky.social
aes2.org
2025 AES International Conference on Artificial Intelligence and Machine Learning for Audio Call for Contributions - AES
Submission Deadline: May 3, 2024
063
Reposted by Pedro Sarmento
Scott H. Hawley @drscotthawley.bsky.social · 23/01/2025
🎉 Follow-up: Thrilled to share that this tutorial has been accepted to #ICLR2025 in the blog posts track!
1161
Reposted by Pedro Sarmento
arXiv Sound @arxiv-sound.bsky.social · 20/01/2025
AI-generated music detection achieved 99.8% accuracy using classifiers trained on real and artificial music. No details on methods or dataset size are provided.
arxiv.org
AI-Generated Music Detection and its Challenges
Darius Afchar, Gabriel Meseguer-Brocal, Romain Hennequin
082
Reposted by Pedro Sarmento
Stefan Lattner @stefanlattner.bsky.social · 20/01/2025
🎶✨ New Paper Announcement! ✨🎶 We present "Improving Musical Accompaniment Co-creation via Diffusion Transformers" 🎹🎸—a study advancing our Diff-A-Riff stem generator through improved quality, efficiency, and control. 📜Read the full paper here: arxiv.org/pdf/2410.23005 🧵👇
arxiv.org
372
Reposted by Pedro Sarmento
Andrew McPherson @apmcpherson.bsky.social · 19/01/2025
First Bsky post, first lab paper of 2025! "On mapping as a technoscientific practice in digital musical instruments" -- a dive on the history and critical implications of mapping theory, with speculation on possible futures. Forthcoming in JNMR: instrumentslab.org/data/andrew/...
1152
Pedro Sarmento @umpedronosapato.bsky.social · 16/01/2025
Let's go 🎸
040
Reposted by Pedro Sarmento
Stephen Roddy @stephenroddy.bandcamp.com · 16/01/2025
Russolo’s intonarumori are in the Guardian. We are so back. www.theguardian.com/music/2025/j...
theguardian.com
Play that funky noise intoner! The rumblers, gurglers and howlers of the world’s strangest orchestra
With their funnel-like hooters and cupboard-like shapes, they looked bizarre, sounded wild and left audiences baffled. Now, more than 100 years on, Luigi Russolo’s orchestra of futurist machine instru...
0195
Reposted by Pedro Sarmento
Marco Comunità @mcomunita.bsky.social · 16/01/2025
Help us with our research to: ☝️ - Develop a perceptual similarity metric for audio effects ✌️ - Advance the state of the art in audio effects modelling Take our listening test (<15min): mcomunita.github.io/mushra-front... Use: 💻 + 🎧 🙏
011
Reposted by Pedro Sarmento
Stefan Lattner @stefanlattner.bsky.social · 15/01/2025
🧑‍🎓 Our #ISMIR Conference Tutorial "Deep Learning 101 for Audio-based MIR" provides a broad introduction to music audio processing, analysis, and generation. 📘 The book and jupyter notebooks: geoffroypeeters.github.io/deeplearning... 🎥 The recording of the tutorial: us02web.zoom.us/rec/share/Qz...
geoffroypeeters.github.io
Deep Learning 101 for Audio-based MIR — Deep Learning 101 for Audio-based MIR
062
Pedro Sarmento @umpedronosapato.bsky.social · 14/01/2025
👀
100
Reposted by Pedro Sarmento
Stefan Lattner @stefanlattner.bsky.social · 14/01/2025
😃 Accepted #ICASSP papers of Sony CSL Music Team: Accompaniment Prompt Adherence: A Measure for Evaluating Music Accompaniment Systems M. Grachten, J. Nistal Estimating Musical Surprisal in Audio M. Bjare, G. Cantisani, S. Lattner and G. Widmer
281
Reposted by Pedro Sarmento
Jordi Pons @jordiponsdotme.bsky.social · 10/01/2025
Weights are out! 🤗 Tokenizing 16kHz speech at very low bitrates. Inference code: github.com/Stability-AI... Model code: github.com/Stability-AI... Model weights: huggingface.co/stabilityai/... arXiv: arxiv.org/abs/2411.19842 Audio demos: stability-ai.github.io/stable-codec...
0204
Reposted by Pedro Sarmento
Oriol (Uri) Nieto @urinieto.bsky.social · 13/01/2025
By far, one of the best things of 2024: www.youtube.com/watch?v=5hTM...
youtube.com
Gojira - Mea Culpa (Ah! Ça ira!) [OFFICIAL VIDEO]
YouTube video by Gojira
041
Pedro Sarmento @umpedronosapato.bsky.social · 13/01/2025
This is such an inspirational insight from Jorge Luis Borges 🙇‍♂️ super relevant today in our quest for more and more and more data
011
Reposted by Pedro Sarmento
Alexandre Défossez @honualx.bsky.social · 13/01/2025
We just released the Helium-1 model , a 2B multi-lingual LLM which @exgrv.bsky.social and @lmazare.bsky.social have been crafting for us! Best model so far under 2.17B params on multi-lingual benchmarks 🇬🇧🇮🇹🇪🇸🇵🇹🇫🇷🇩🇪 On HF, under CC-BY licence: huggingface.co/kyutai/heliu...
0258
Reposted by Pedro Sarmento
Jesse Engel @jesseengel.bsky.social · 13/01/2025
What an astonishingly tone deaf take. He misses the entire point of creative practice. Accessibility and mastery are not opposed, they're two ends of the same personal creative journey.
161
Reposted by Pedro Sarmento
Vincent Lostanlen @lostanlen.bsky.social · 10/01/2025
I am organizing a session on "Advancements in Bird Communication Studies" at Forum Acusticum / Euronoise, to be held in Málaga on June 23–26, 2025. Please submit your abstract (max. 200 words) onto the conference portal before January 19th www.fa-euronoise2025.org/abstract-sub...
fa-euronoise2025.org
Abstract submission - Forum Acusticum Euronoise 2025
Forum Acusticum Euronoise 2025
062
Reposted by Pedro Sarmento
keunwoochoi.bsky.social @keunwoochoi.bsky.social · 09/01/2025
fixed - ismir2024 papers are accessible now! along with the reviews too, sometimes. ismir2024program.ismir.net
ismir2024program.ismir.net
ISMIR 2024: Schedule
071
Reposted by Pedro Sarmento
Ethan Mollick @emollick.bsky.social · 09/01/2025
Interesting attempt to build an AI agent-based research assistant to automate machine learning paper writing by acting as PhDs, post-docs, & professors working in a typical lab. It doesn't autonomously produce high-level work but it looks promising as a copilot for researchers cutting cost & effort.
2467
Reposted by Pedro Sarmento
QMUL School of Electronic Engineering and Computer Science @qmuleecs.bsky.social · 09/01/2025
🤖 If you missed Prof. Shalom Lappin's insightful lectures series on the core ideas of his forthcoming book, 'Understanding #AI: Neither Catastrophe nor Redemption', you can catch up and watch the recordings on our Youtube channel: www.youtube.com/@QMEECS/videos
021
Pedro Sarmento @umpedronosapato.bsky.social · 17/12/2024
Awesome DMRN @c4dm.bsky.social workshop today, with a great keynote by titan @stefanlattner.bsky.social and looooads of research around guitar 🤘🎸 @jackjamesloth.bsky.social
093
Pedro Sarmento @umpedronosapato.bsky.social · 12/12/2024
another banger work by titan @hugofloresgarcia.bsky.social 🤘
020
Reposted by Pedro Sarmento
Anil Ananthaswamy @anilananth.bsky.social · 11/12/2024
The latest Science-in-Parallel episode dropped, in which I talk of this epochal moment in human history (the coming of LLMs), the 2024 NobelPrize for Hinton and Hopfield, and the history of neural networks, besides the writing of WHY MACHINES LEARN. scienceinparallel.org/2024/12/anil...
scienceinparallel.org
Anil Ananthaswamy: AI's Nobel Moment - Science in Parallel
2024 was artificial intelligence’s Nobel Prize year with the physics and chemistry prizes recognizing the underpinnings and application of these […]
0112
Reposted by Pedro Sarmento
arXiv Sound @arxiv-sound.bsky.social · 10/12/2024
A deep learning pipeline uses spectrogram masking and the MuseScore API to separate instrument stems from music audio, convert them to MIDI, and transcribe them into sheet music.
arxiv.org
Source Separation Automatic Transcription for Music
Bradford Derby, Lucas Dunker, Samarth Galchar, Shashank Jarmale, Akash Setti
011
Reposted by Pedro Sarmento
DCASE Challenge @dcase-challenge.bsky.social · 10/12/2024
The tasks for DCASE challenge 2025 have been announced. dcase.community/articles/cha... Stay tuned for more details.
dcase.community
Challenge tasks for DCASE2025 - DCASE
The DCASE Steering Group has reviewed the task proposals...
074
Reposted by Pedro Sarmento
Anil Ananthaswamy @anilananth.bsky.social · 07/12/2024
What? Linear algebra and calculus and machine learning for the holidays! Might a math-y book be a good gift for the holidays? I hope so :-) “A masterpiece.”-Geoff Hinton “A masterful work.”-Melanie Mitchell US www.penguinrandomhouse.com/books/677608... UK www.penguin.co.uk/books/446849...
penguinrandomhouse.com
Why Machines Learn by Anil Ananthaswamy: 9780593185742 | PenguinRandomHouse.com: Books
A rich, narrative explanation of the mathematics that has brought us machine learning and the ongoing explosion of artificial intelligence Machine learning systems are making life-altering decisions...
0132
Reposted by Pedro Sarmento
roser @roserbatlleroca.bsky.social · 03/12/2024
📢 Call to all #MIRCommunity! Have you ever wondered what an open model means? Help us shape the definition of open models in generative AI for music by taking our survey — just 10 minutes! 👉 forms.gle/Z48t6HPBXwWC3r… thank you 💫
forms.gle
032
Pedro Sarmento @umpedronosapato.bsky.social · 03/12/2024
join us 🤘
110
Reposted by Pedro Sarmento
Sander Dieleman @sedielem.bsky.social · 02/12/2024
There is a lot of great writing on flow matching out there all of a sudden! This post clarifies the connection with diffusion models -- they are essentially two different ways to describe the same class of models.
0273
Reposted by Pedro Sarmento
Scott H. Hawley @drscotthawley.bsky.social · 02/12/2024
Thanks! BTW for anyone seeing this: I made my repo public (though still WIP) & would welcome feedback to improve the results: github.com/drscotthawle... e.g. 1. I haven't even added attention yet, and 2. I'm not sure the "GAN" part is really learning. Sample input/recon after 6 hours:
022