Reposted by Alberto Hernández MarcosCésar Astudillo @cesarastudillo.bsky.social · 19/07/2025Este artículo me ha parecido muy provocador y me ha abierto muchas dudas. Quizá un feminismo plural y dialogante pueda tener conversaciones incómodas y constructivas con una fracción de los hombres objetivo del discurso manosférico, pero temo que para los propios manosféricos ya sea tarde para eso 151
Reposted by Alberto Hernández MarcosLuis Saiz @lsaiz.bsky.social · 21/06/2025Rubén Santamarta ha venido publicado análisis fundamentados sobre el apagón Ahora está pidiendo logs de los que tengáis fotovoltaica conectada a red www.linkedin.com/posts/rubens... cc/ @mjelectriz.bsky.social @revenergetica.bsky.social @todoselectricos.bsky.social @pacovalverde.bsky.sociallinkedin.comMe gustaría apelar a vuestra colaboración para poder profundizar en el análisis ciber-físico del papel de los inversores solares de autoconsumo en el apagón. | Ruben SantamartaMe gustaría apelar a vuestra colaboración para poder profundizar en el análisis ciber-físico del papel de los inversores solares de autoconsumo en el apagón. Los que me seguís ya sabéis que he estado... 3116
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 15/06/2025A big AI question is why, as LLMs get bigger, their values seem to increasingly converge on the same preferences, this holds for Musk’s Grok & China’s DeepSeek, too. “These findings suggest that value systems emerge in LLMs in a meaningful sense, with broad implications” arxiv.org/abs/2502.08640 2418722
Reposted by Alberto Hernández MarcosCarl T. Bergstrom @carlbergstrom.com · 09/06/2025Ok, time for a short thread about this paper. My sense over the past six months or so is that chain-of-thought prompting as used in e.g. ChatGPT o.3 improves substantially upon previous systems such as ChatGPT 4.o, at least for certain tasks. But how revolutionary is it? 1628892
Reposted by Alberto Hernández MarcosDean Baker @deanbaker13.bsky.social · 27/05/2025I strongly second Krugman's "letter to Europe." The EU should tell Trump to take his tariff and shove it. Hitting U.S. consumers with a huge tax increase is not smart policy, but as a reality TV show star, what does Trump know about economics? paulkrugman.substack.com/p/a-letter-t...paulkrugman.substack.comA Letter to EuropeYou’re stronger than you think. Act like it. 17685180
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 22/05/2025Individuals keep self-reporting huge gains in productivity from AI & controlled experiments in many industries keep finding these boosts are real, yet most firms are not seeing big effects. Why? Because gaining from AI requires organizational innovation. www.oneusefulthing.org/p/making-ai-...oneusefulthing.orgMaking AI Work: Leadership, Lab, and CrowdA formula for AI in companies 25512
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 20/05/2025Big: The final version of a randomized, controlled World Bank study finds using a GPT-4 tutor with teacher guidance in a six week afterschool program in Nigeria had "more than twice the effect of some of the most effective interventions in education" ("equating to 1.5 to 2 years" of standard school) 1216833
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 16/05/2025I wish these skeptical AI articles (this is from the NYTimes) would actually grapple with the growing body of research that AI can do original research & perform key unstructured tasks across the spectrum of high-end white collar employment. AI criticism is important, but it should be clear-eyed. 5527
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 10/05/2025A common question is "can an AI make money?" This benchmark, where AIs run a simulated vending machine over time, suggests yes, with an important caveat On average, Claude 3.5 & o3-mini beat a human, but are high in variance & fail at random times for complex reasons. andonlabs.com/evals/vendin... 86615
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 08/05/2025"Our findings demonstrate that reasoning models improve not only the clarity, organization, and professionalism of legal work but also the depth & rigor of legal analysis itself." Law students using o1-preview had the quality of their work on most tasks increase (up to 28%) & time savings of 12-28% 1545
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 06/05/2025I just don’t see signs of a major increase in hallucination rates for recent models, or for reasoners overall, in the data. It seems like some models do better than others, but many of the recent models have the lowest hallucination rates. 38012
Reposted by Alberto Hernández MarcosMelanie Mitchell @melaniemitchell.bsky.social · 06/05/2025I'm looking forward to this event next week in Amsterdam! 1253
Reposted by Alberto Hernández Marcosruggsea @ruggsea.eurosky.social · 07/05/2025Reposting from the other site to spread it here: apparently, thinking that RLHF irons out creativity in LLMs is now corroborated by this paper arxiv.org/pdf/2505.00047 1173
Reposted by Alberto Hernández MarcosKeezy Young🌼 @keezyyoung.bsky.social · 07/05/2025the specific laughter of a little kid who is being taken on a ride of some kind (tricycle, plastic car, thrown up in the air by a dad etc) is one of the most precious and valuable things you'll ever hear 11568
Reposted by Alberto Hernández MarcosMelanie Mitchell @melaniemitchell.bsky.social · 02/05/2025Karpathy: We have reached "jagged Intelligence" Mollick: We have reached "jagged AGI" Next up: "jagged consciousness"? www.oneusefulthing.org/p/on-jagged-...oneusefulthing.orgOn Jagged AGI: o3, Gemini 2.5, and everything afterNew models and new thresholds 6426
Reposted by Alberto Hernández MarcosCésar Astudillo @cesarastudillo.bsky.social · 29/04/2025A quienes teníais pensado acudir: por desgracia habrá que aplazar el evento porque la UCM ha suspendido todas las actividades. Se establecerá una nueva fecha a corto plazo y por supuesto os la contaré. 041
Reposted by Alberto Hernández MarcosGus @gusthema.bsky.social · 28/04/2025Gemma 3 are just amazing models! but what if you want to manipulate it's internal activations to understand how it does its text generation? Sascha Rothe is here to teach you how! Great insights for anyone curious about the inner workings of LLMs! www.youtube.com/watch?v=JTUs...youtube.comInside Gemma 3: Modifying the output through activation hackingYouTube video by Google for Developers 1104
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 28/04/2025👀Today’s AIs are already hyper persuasive. A controversial study where LLMs tried to persuade users on Reddit found that: “Notably, all our treatments surpass human performance substantially, achieving persuasive rates between three and six times higher than the human baseline.” 911015
Reposted by Alberto Hernández MarcosPíxel Sonoro @pixelsonoro.bsky.social · 28/04/2025🎵🎶¡Mañana celebramos este evento en homenaje a la obra de @cesarastudillo.bsky.social y @riskwood.bsky.social donde profundizaremos en la creación musical durante la "edad de oro" ✅Exposiciones ✅Interpretación de arreglos orquestales de algunas de sus piezas ¿Cómo puedes seguirlo? ⬇️⬇️ 155
Reposted by Alberto Hernández MarcosPhillip Carter @phillipcarter.dev · 10/04/2025I have yet to see an "intro to AI" video as comprehensive but also approachable as Andrej Karpathy's 1hr overview. I maintain that anyone who watches this (and pays attention the whole time) will come away with an intuitive understanding of how to use LLMs www.youtube.com/watch?v=zjkB...youtube.com[1hr Talk] Intro to Large Language ModelsYouTube video by Andrej Karpathy 1172
Reposted by Alberto Hernández MarcosKate Beaton @katebeaton.bsky.social · 12/04/2025Richard Serra lived seasonally in our village so we went to the Reina Sofia to see their room of his work, but the children had no love for grand abstract minimalist rectangles, and we were forced to depart after much whining, it was a quixotic quest anyway 82523
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 03/04/2025If you wanted to see how little attention folks are paying to the possibility of AGI (however defined) no matter how much the labs publicly discuss it, here is an official course from Google Deepmind whose first session is "we are on a path to superhuman capabilities" It has less than 1,000 views. 2214416
Alberto Hernández Marcos @alberto-h.bsky.social · 02/04/2025And this applies to so many other approaches, like building-my-own-RAG... 010
Reposted by Alberto Hernández MarcosSerge Belongie @serge.belongie.com · 30/03/2025Would you present your next NeurIPS paper in Europe instead of traveling to San Diego (US) if this was an option? Søren Hauberg (DTU) and I would love to hear the answer through this poll: (1/6)docs.google.comNeurIPS participation in EuropeWe seek to understand if there is interest in being able to attend NeurIPS in Europe, i.e. without travelling to San Diego, US. In the following, assume that it is possible to present accepted papers ... 6279161
Reposted by Alberto Hernández MarcosPete Marcus @petemarcus.bsky.social · 27/03/2025Useful chart to understand AI agents (by @mmitchell.bsky.social) www.technologyreview.com/2025/03/24/1... 0197
Reposted by Alberto Hernández MarcosMelanie Mitchell @melaniemitchell.bsky.social · 20/03/2025In my latest column for Science magazine, I discuss recent AI "reasoning" models -- how it works, to what extent it captures "genuine" reasoning processes, and what's needed to answer such questions. www.science.org/doi/10.1126/...science.orgArtificial intelligence learns to reasonJulia has two sisters and one brother. How many sisters does her brother Martin have?Solving this tiny puzzle requires a bit of thinking. You might mentally picture the family of three girls and one b... 715759
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 20/03/2025This trend seems to be growing here. False comfort that AI doesn’t work or that it isn’t getting better is pervasive on Blue Sky. As a result, people who could add important points of view to current discussions on the meaning & use of AI instead try to believe they don’t have to think about it. 2224641
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 04/03/2025Not to be a broken record, but AI critics who insist that AI "doesn't work" and is going to just disappear are misleading - that just isn't true, as controlled studies like this one show. There are many issues with AI & many things that need critique, but pretending it is going away is not helpful. 921732
Alberto Hernández Marcos @alberto-h.bsky.social · 19/03/2025Can #AI learn and produce its own emotions, like natural ones? 🤖❤️ Meet LOVE (Latest Observed Values Encoding), a generic self-learning emotional framework for machines. Paper in Nature - Scientific Reports (open access): nature.com/articles/s41598-024-72817-x See how it works! 🧵⬇️nature.comA generic self-learning emotional framework for machines - Scientific ReportsScientific Reports - A generic self-learning emotional framework for machines 120
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 15/03/2025“Gemini, remove the squid from this picture from the movie All Quiet on the Western Front” “But there is no squid in the original image“ “Remove the squid” “I will visually emphasize the absolute absence of a squid” “Still might be squid somewhere” “How about now?” “Well…” 6614
Reposted by Alberto Hernández MarcosLuis Saiz @lsaiz.bsky.social · 10/03/2025Long, but profound and worthy 021
Reposted by Alberto Hernández MarcosJaime Gómez-Obregón @gomezobregon.com · 09/03/2025Para todos los interesados en perder el tiempo en algo absurdo, os dejo este curso de Flash que promueve el Servicio Público de Empleo de Castilla y León y financia el Ministerio de Educación. 125532
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 07/03/2025I have mixed feelings about the phrase "vibe coding" but I really feel the appeal of conjuring something to life with words. I asked Veo2 for: "a sweeping brutalist arch looming over a suburban town as a man & woman, their arms completely covered with flowers, bike through the flooded streets" 3785
Reposted by Alberto Hernández MarcosTed Underwood @tedunderwood.com · 07/03/2025Building Bluesky has to have been the most stressful and gratifying job. All the anxiety of throwing a party—except several million people show up before you’ve built the house. You have to build it with them inside. And no pressure, but if it works, we’ll have a new desperately needed institution. 2526
Reposted by Alberto Hernández MarcosCésar Astudillo @cesarastudillo.bsky.social · 05/03/2025La teoría subjetiva del valor explica un fenómeno frecuente: algo puede haberte costado muy poco, y tener muchísimo valor para alguien. Eso es una oportunidad para generar grandes beneficios. Leer "beneficios" en clave estrictamente contable o de modo más general ya es elección tuya. 3182
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 22/02/2025As AI models grow in size, their values seem to increasingly converge on the same preferences, whether the model is made by OpenAI or Musk’s X or China’s DeepSeek Not everyone will like resulting AI preferences, so value engineering is likely to become a topic of discussion arxiv.org/abs/2502.08640 47712
Alberto Hernández Marcos @alberto-h.bsky.social · 18/02/2025@karpathy.bsky.social We're seriously missing you in Bluesky... 010
Reposted by Alberto Hernández MarcosEugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 14/02/2025Model-free deep RL algorithms like NFSP, PSRO, ESCHER, & R-NaD are tailor-made for games with hidden information (e.g. poker). We performed the largest-ever comparison of these algorithms. We find that they do not outperform generic policy gradient methods, such as PPO. arxiv.org/abs/2502.08938 1/N 39321
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 11/02/2025There is a lot going on in this paper, but it shows that no matter their maker's country (China or US) or politics (Musk's Grok vs. OpenAI), models seem to converge to the same values, often with some fairly shocking results. Those values might be steered in the future, but by who is a big question. 48920
Reposted by Alberto Hernández MarcosMargaret Mitchell @mmitchell.bsky.social · 10/02/2025⚫⚪ It's coming...SHADES. ⚪⚫ The first ever resource of multilingual, multicultural, and multigeographical stereotypes, built to support nuanced LLM evaluation and bias mitigation. We have been working on this around the world for almost **4 years** and I am thrilled to share it with you all soon. 612723
Reposted by Alberto Hernández MarcosMelanie Mitchell @melaniemitchell.bsky.social · 30/01/2025Very good (technical) explainer answering "How has DeepSeek improved the Transformer architecture?". Aimed at readers already familiar with Transformers. epoch.ai/gradient-upd...epoch.aiHow has DeepSeek improved the Transformer architecture?This Gradient Updates issue goes over the major changes that went into DeepSeek’s most recent model. 627763
Reposted by Alberto Hernández MarcosEd Hawkins @edhawkins.org · 16/01/2025As the climate data for 2024 continues to arrive, I've updated my graphic showing changes in various climate indicators for the recent past, and the last 2000 years. Most notable change is a big jump for tropospheric temperatures (TLT) in 2024. [Sea level and land humidity data not yet available.] 11324129
Reposted by Alberto Hernández MarcosEthan Mollick @emollick.bsky.social · 11/01/2025The British infantry at Waterloo during a French cavalry charge, the redcoats are all wearing rubber ducks on their heads. Veo2 showing history as it was. 4865
Reposted by Alberto Hernández MarcosThe Atlantic @theatlantic.com · 11/01/2025“In one of the most astonishing political transformations in the history of democracy,” Hitler destroyed “a constitutional republic through constitutional means,” writes Timothy W. Ryback. Read about how Hitler overcame democracy in just 53 days:theatlantic.comHow Hitler Dismantled a Democracy in 53 DaysHe used the constitution to shatter the constitution. 52511260
Reposted by Alberto Hernández MarcosJeff Dean @jeffdean.bsky.social · 05/01/2025This guide to starter packs has a wealth of different lists of people with various interests: blueskydirectory.com/starter-pack...blueskydirectory.comBluesky Starter Packs - Bluesky DirectoryBrowse a list of Bluesky Starter Packs. Discover and connect with your community on Bluesky 2337
Reposted by Alberto Hernández MarcosCésar Astudillo @cesarastudillo.bsky.social · 05/01/2025¡Repost, porfa! INVESTIGACIÓN FABLABS ¿Has gestionado un FabLab? ¿Tienes experiencia como usuaria o usuario de uno? Tengo unas alumnas de master que quieren hacer su TFG sobre el tema. ¿Puedo ponerte en contacto con ellas para una entrevista de 30 min? Manda DM 1714
Reposted by Alberto Hernández MarcosPeyman Milanfar @docmilanfar.bsky.social · 05/01/2025the five horsemen of the apocalypse 1193
Reposted by Alberto Hernández MarcosGurwinder @gurwinder.bsky.social · 01/01/202525 USEFUL CONCEPTS TO HELP YOU GET THROUGH 2025 Thread: 58434