Sign in

Alberto Hernández Marcos

@alberto-h.bsky.social
97 followers 116 following 68 posts

AI, pixels and R&R. Head of GenAI Lab @BBVA. PhD (University of Granada) on emotion-driven Reinforcement Learning 🤖 + ❤️ Opinions are my own

PostsRepliesMedia
Reposted by Alberto Hernández Marcos
César Astudillo @cesarastudillo.bsky.social · 19/07/2025
Este artículo me ha parecido muy provocador y me ha abierto muchas dudas. Quizá un feminismo plural y dialogante pueda tener conversaciones incómodas y constructivas con una fracción de los hombres objetivo del discurso manosférico, pero temo que para los propios manosféricos ya sea tarde para eso
151
Reposted by Alberto Hernández Marcos
Luis Saiz @lsaiz.bsky.social · 21/06/2025
Rubén Santamarta ha venido publicado análisis fundamentados sobre el apagón Ahora está pidiendo logs de los que tengáis fotovoltaica conectada a red www.linkedin.com/posts/rubens... cc/ @mjelectriz.bsky.social @revenergetica.bsky.social @todoselectricos.bsky.social @pacovalverde.bsky.social
linkedin.com
Me gustaría apelar a vuestra colaboración para poder profundizar en el análisis ciber-físico del papel de los inversores solares de autoconsumo en el apagón. | Ruben Santamarta
Me gustaría apelar a vuestra colaboración para poder profundizar en el análisis ciber-físico del papel de los inversores solares de autoconsumo en el apagón. Los que me seguís ya sabéis que he estado...
3116
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 15/06/2025
A big AI question is why, as LLMs get bigger, their values seem to increasingly converge on the same preferences, this holds for Musk’s Grok & China’s DeepSeek, too. “These findings suggest that value systems emerge in LLMs in a meaningful sense, with broad implications” arxiv.org/abs/2502.08640
2418722
Reposted by Alberto Hernández Marcos
Carl T. Bergstrom @carlbergstrom.com · 09/06/2025
Ok, time for a short thread about this paper. My sense over the past six months or so is that chain-of-thought prompting as used in e.g. ChatGPT o.3 improves substantially upon previous systems such as ChatGPT 4.o, at least for certain tasks. But how revolutionary is it?
1628892
Reposted by Alberto Hernández Marcos
Dean Baker @deanbaker13.bsky.social · 27/05/2025
I strongly second Krugman's "letter to Europe." The EU should tell Trump to take his tariff and shove it. Hitting U.S. consumers with a huge tax increase is not smart policy, but as a reality TV show star, what does Trump know about economics? paulkrugman.substack.com/p/a-letter-t...
paulkrugman.substack.com
A Letter to Europe
You’re stronger than you think. Act like it.
17685180
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 22/05/2025
Individuals keep self-reporting huge gains in productivity from AI & controlled experiments in many industries keep finding these boosts are real, yet most firms are not seeing big effects. Why? Because gaining from AI requires organizational innovation. www.oneusefulthing.org/p/making-ai-...
oneusefulthing.org
Making AI Work: Leadership, Lab, and Crowd
A formula for AI in companies
25512
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 20/05/2025
Big: The final version of a randomized, controlled World Bank study finds using a GPT-4 tutor with teacher guidance in a six week afterschool program in Nigeria had "more than twice the effect of some of the most effective interventions in education" ("equating to 1.5 to 2 years" of standard school)
1216833
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 16/05/2025
I wish these skeptical AI articles (this is from the NYTimes) would actually grapple with the growing body of research that AI can do original research & perform key unstructured tasks across the spectrum of high-end white collar employment. AI criticism is important, but it should be clear-eyed.
5527
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 10/05/2025
A common question is "can an AI make money?" This benchmark, where AIs run a simulated vending machine over time, suggests yes, with an important caveat On average, Claude 3.5 & o3-mini beat a human, but are high in variance & fail at random times for complex reasons. andonlabs.com/evals/vendin...
86615
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 08/05/2025
"Our findings demonstrate that reasoning models improve not only the clarity, organization, and professionalism of legal work but also the depth & rigor of legal analysis itself." Law students using o1-preview had the quality of their work on most tasks increase (up to 28%) & time savings of 12-28%
1545
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 06/05/2025
I just don’t see signs of a major increase in hallucination rates for recent models, or for reasoners overall, in the data. It seems like some models do better than others, but many of the recent models have the lowest hallucination rates.
38012
Reposted by Alberto Hernández Marcos
Melanie Mitchell @melaniemitchell.bsky.social · 06/05/2025
I'm looking forward to this event next week in Amsterdam!
1253
Reposted by Alberto Hernández Marcos
ruggsea @ruggsea.eurosky.social · 07/05/2025
Reposting from the other site to spread it here: apparently, thinking that RLHF irons out creativity in LLMs is now corroborated by this paper arxiv.org/pdf/2505.00047
1173
Reposted by Alberto Hernández Marcos
Keezy Young🌼 @keezyyoung.bsky.social · 07/05/2025
the specific laughter of a little kid who is being taken on a ride of some kind (tricycle, plastic car, thrown up in the air by a dad etc) is one of the most precious and valuable things you'll ever hear
11568
Reposted by Alberto Hernández Marcos
Melanie Mitchell @melaniemitchell.bsky.social · 02/05/2025
Karpathy: We have reached "jagged Intelligence" Mollick: We have reached "jagged AGI" Next up: "jagged consciousness"? www.oneusefulthing.org/p/on-jagged-...
oneusefulthing.org
On Jagged AGI: o3, Gemini 2.5, and everything after
New models and new thresholds
6426
Reposted by Alberto Hernández Marcos
César Astudillo @cesarastudillo.bsky.social · 29/04/2025
A quienes teníais pensado acudir: por desgracia habrá que aplazar el evento porque la UCM ha suspendido todas las actividades. Se establecerá una nueva fecha a corto plazo y por supuesto os la contaré.
041
Reposted by Alberto Hernández Marcos
Gus @gusthema.bsky.social · 28/04/2025
Gemma 3 are just amazing models! but what if you want to manipulate it's internal activations to understand how it does its text generation? Sascha Rothe is here to teach you how! Great insights for anyone curious about the inner workings of LLMs! www.youtube.com/watch?v=JTUs...
youtube.com
Inside Gemma 3: Modifying the output through activation hacking
YouTube video by Google for Developers
1104
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 28/04/2025
👀Today’s AIs are already hyper persuasive. A controversial study where LLMs tried to persuade users on Reddit found that: “Notably, all our treatments surpass human performance substantially, achieving persuasive rates between three and six times higher than the human baseline.”
911015
Reposted by Alberto Hernández Marcos
Píxel Sonoro @pixelsonoro.bsky.social · 28/04/2025
🎵🎶¡Mañana celebramos este evento en homenaje a la obra de @cesarastudillo.bsky.social y @riskwood.bsky.social donde profundizaremos en la creación musical durante la "edad de oro" ✅Exposiciones ✅Interpretación de arreglos orquestales de algunas de sus piezas ¿Cómo puedes seguirlo? ⬇️⬇️
155
Reposted by Alberto Hernández Marcos
Phillip Carter @phillipcarter.dev · 10/04/2025
I have yet to see an "intro to AI" video as comprehensive but also approachable as Andrej Karpathy's 1hr overview. I maintain that anyone who watches this (and pays attention the whole time) will come away with an intuitive understanding of how to use LLMs www.youtube.com/watch?v=zjkB...
youtube.com
[1hr Talk] Intro to Large Language Models
YouTube video by Andrej Karpathy
1172
Reposted by Alberto Hernández Marcos
Kate Beaton @katebeaton.bsky.social · 12/04/2025
Richard Serra lived seasonally in our village so we went to the Reina Sofia to see their room of his work, but the children had no love for grand abstract minimalist rectangles, and we were forced to depart after much whining, it was a quixotic quest anyway
Richard Serra sculptures A person standing in front of a Richard Serra artwork
82523
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 03/04/2025
If you wanted to see how little attention folks are paying to the possibility of AGI (however defined) no matter how much the labs publicly discuss it, here is an official course from Google Deepmind whose first session is "we are on a path to superhuman capabilities" It has less than 1,000 views.
2214416
Alberto Hernández Marcos @alberto-h.bsky.social · 02/04/2025
And this applies to so many other approaches, like building-my-own-RAG...
010
Reposted by Alberto Hernández Marcos
Serge Belongie @serge.belongie.com · 30/03/2025
Would you present your next NeurIPS paper in Europe instead of traveling to San Diego (US) if this was an option? Søren Hauberg (DTU) and I would love to hear the answer through this poll: (1/6)
docs.google.com
NeurIPS participation in Europe
We seek to understand if there is interest in being able to attend NeurIPS in Europe, i.e. without travelling to San Diego, US. In the following, assume that it is possible to present accepted papers ...
6279161
Reposted by Alberto Hernández Marcos
Pete Marcus @petemarcus.bsky.social · 27/03/2025
Useful chart to understand AI agents (by @mmitchell.bsky.social) www.technologyreview.com/2025/03/24/1...
0197
Reposted by Alberto Hernández Marcos
Paul Graham @paulgbot.bsky.social · 15/03/2025
tweet screenshot
043
Reposted by Alberto Hernández Marcos
Melanie Mitchell @melaniemitchell.bsky.social · 20/03/2025
In my latest column for Science magazine, I discuss recent AI "reasoning" models -- how it works, to what extent it captures "genuine" reasoning processes, and what's needed to answer such questions. www.science.org/doi/10.1126/...
science.org
Artificial intelligence learns to reason
Julia has two sisters and one brother. How many sisters does her brother Martin have?Solving this tiny puzzle requires a bit of thinking. You might mentally picture the family of three girls and one b...
715759
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 20/03/2025
This trend seems to be growing here. False comfort that AI doesn’t work or that it isn’t getting better is pervasive on Blue Sky. As a result, people who could add important points of view to current discussions on the meaning & use of AI instead try to believe they don’t have to think about it.
2224641
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 04/03/2025
Not to be a broken record, but AI critics who insist that AI "doesn't work" and is going to just disappear are misleading - that just isn't true, as controlled studies like this one show. There are many issues with AI & many things that need critique, but pretending it is going away is not helpful.
921732
Alberto Hernández Marcos @alberto-h.bsky.social · 19/03/2025
Can #AI learn and produce its own emotions, like natural ones? 🤖❤️ Meet LOVE (Latest Observed Values Encoding), a generic self-learning emotional framework for machines. Paper in Nature - Scientific Reports (open access): nature.com/articles/s41598-024-72817-x See how it works! 🧵⬇️
nature.com
A generic self-learning emotional framework for machines - Scientific Reports
Scientific Reports - A generic self-learning emotional framework for machines
120
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 15/03/2025
“Gemini, remove the squid from this picture from the movie All Quiet on the Western Front” “But there is no squid in the original image“ “Remove the squid” “I will visually emphasize the absolute absence of a squid” “Still might be squid somewhere” “How about now?” “Well…”
6614
Reposted by Alberto Hernández Marcos
Luis Saiz @lsaiz.bsky.social · 10/03/2025
Long, but profound and worthy
021
Reposted by Alberto Hernández Marcos
Jaime Gómez-Obregón @gomezobregon.com · 09/03/2025
Para todos los interesados en perder el tiempo en algo absurdo, os dejo este curso de Flash que promueve el Servicio Público de Empleo de Castilla y León y financia el Ministerio de Educación.
125532
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 07/03/2025
I have mixed feelings about the phrase "vibe coding" but I really feel the appeal of conjuring something to life with words. I asked Veo2 for: "a sweeping brutalist arch looming over a suburban town as a man & woman, their arms completely covered with flowers, bike through the flooded streets"
3785
Reposted by Alberto Hernández Marcos
Ted Underwood @tedunderwood.com · 07/03/2025
Building Bluesky has to have been the most stressful and gratifying job. All the anxiety of throwing a party—except several million people show up before you’ve built the house. You have to build it with them inside. And no pressure, but if it works, we’ll have a new desperately needed institution.
2526
Reposted by Alberto Hernández Marcos
César Astudillo @cesarastudillo.bsky.social · 05/03/2025
La teoría subjetiva del valor explica un fenómeno frecuente: algo puede haberte costado muy poco, y tener muchísimo valor para alguien. Eso es una oportunidad para generar grandes beneficios. Leer "beneficios" en clave estrictamente contable o de modo más general ya es elección tuya.
3182
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 22/02/2025
As AI models grow in size, their values seem to increasingly converge on the same preferences, whether the model is made by OpenAI or Musk’s X or China’s DeepSeek Not everyone will like resulting AI preferences, so value engineering is likely to become a topic of discussion arxiv.org/abs/2502.08640
47712
Alberto Hernández Marcos @alberto-h.bsky.social · 18/02/2025
@karpathy.bsky.social We're seriously missing you in Bluesky...
010
Reposted by Alberto Hernández Marcos
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 14/02/2025
Model-free deep RL algorithms like NFSP, PSRO, ESCHER, & R-NaD are tailor-made for games with hidden information (e.g. poker). We performed the largest-ever comparison of these algorithms. We find that they do not outperform generic policy gradient methods, such as PPO. arxiv.org/abs/2502.08938 1/N
39321
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 11/02/2025
There is a lot going on in this paper, but it shows that no matter their maker's country (China or US) or politics (Musk's Grok vs. OpenAI), models seem to converge to the same values, often with some fairly shocking results. Those values might be steered in the future, but by who is a big question.
48920
Reposted by Alberto Hernández Marcos
Margaret Mitchell @mmitchell.bsky.social · 10/02/2025
⚫⚪ It's coming...SHADES. ⚪⚫ The first ever resource of multilingual, multicultural, and multigeographical stereotypes, built to support nuanced LLM evaluation and bias mitigation. We have been working on this around the world for almost **4 years** and I am thrilled to share it with you all soon.
Screenshot of 'SHADES: Towards a Multilingual Assessment of Stereotypes in Large Language Models.'
SHADES is in multiple grey colors (shades).
612723
Reposted by Alberto Hernández Marcos
Melanie Mitchell @melaniemitchell.bsky.social · 30/01/2025
Very good (technical) explainer answering "How has DeepSeek improved the Transformer architecture?". Aimed at readers already familiar with Transformers. epoch.ai/gradient-upd...
epoch.ai
How has DeepSeek improved the Transformer architecture?
This Gradient Updates issue goes over the major changes that went into DeepSeek’s most recent model.
627763
Reposted by Alberto Hernández Marcos
Ed Hawkins @edhawkins.org · 16/01/2025
As the climate data for 2024 continues to arrive, I've updated my graphic showing changes in various climate indicators for the recent past, and the last 2000 years. Most notable change is a big jump for tropospheric temperatures (TLT) in 2024. [Sea level and land humidity data not yet available.]
Various climate indicators for the last 2024 years.
11324129
Reposted by Alberto Hernández Marcos
Ethan Mollick @emollick.bsky.social · 11/01/2025
The British infantry at Waterloo during a French cavalry charge, the redcoats are all wearing rubber ducks on their heads. Veo2 showing history as it was.
4865
Reposted by Alberto Hernández Marcos
The Atlantic @theatlantic.com · 11/01/2025
“In one of the most astonishing political transformations in the history of democracy,” Hitler destroyed “a constitutional republic through constitutional means,” writes Timothy W. Ryback. Read about how Hitler overcame democracy in just 53 days:
theatlantic.com
How Hitler Dismantled a Democracy in 53 Days
He used the constitution to shatter the constitution.
52511260
Reposted by Alberto Hernández Marcos
Jeff Dean @jeffdean.bsky.social · 05/01/2025
This guide to starter packs has a wealth of different lists of people with various interests: blueskydirectory.com/starter-pack...
blueskydirectory.com
Bluesky Starter Packs - Bluesky Directory
Browse a list of Bluesky Starter Packs. Discover and connect with your community on Bluesky
2337
Reposted by Alberto Hernández Marcos
César Astudillo @cesarastudillo.bsky.social · 05/01/2025
¡Repost, porfa! INVESTIGACIÓN FABLABS ¿Has gestionado un FabLab? ¿Tienes experiencia como usuaria o usuario de uno? Tengo unas alumnas de master que quieren hacer su TFG sobre el tema. ¿Puedo ponerte en contacto con ellas para una entrevista de 30 min? Manda DM
1714
Reposted by Alberto Hernández Marcos
Peyman Milanfar @docmilanfar.bsky.social · 05/01/2025
the five horsemen of the apocalypse
1193
Reposted by Alberto Hernández Marcos
Gurwinder @gurwinder.bsky.social · 01/01/2025
25 USEFUL CONCEPTS TO HELP YOU GET THROUGH 2025 Thread:
58434