Hey all, well, I have been wanting to post in LessWrong for a while but I don't have what is worth being put into a post.
I am a 19 years old AI and robotics engineering undergraduate student from Iraq. Well I am in my sophomore year and I have went through a lot of changes in my mapping of reality. I was raised into Islam, I left that initially when I was 11 but fully left it at 14, I think some books I read then triggered it. I went through an intense 3 months derealization experience after that where I completely lost all sense of reality and self, I had trouble sleeping back then lol. Well that passed and I began building my beliefs from scratch, I read so much philosophy from both east and west, I jumped through a lot of beliefs.
I got into university and I have fallen in love with engineering, but seeing how the world is using AI and the negative effects it is having was quite sad for me, though I didn't know of the field of AI safety. I became digital privacy obsessed for a while, just to make sure that big corporations don't collect my data to get used in algorithmic targeting and AI training.
In the meantime, I got into so much experiences, met so much people, had so many c... (read more)
Hello, and thanks @habryka! I’ve been vaguely aware of LW for several years but only today published my first post after reading recent discussion about whether forecasting is ‘worth it.’
I have quite a specific perspective on this - I work in the tech industry as a product strategist, and the main way I do my work is using forecasting and foresight methods to help people make decisions.
I’m hoping this is a place where I can write more regularly; as well as broaden my knowledge of salient ideas in AI progress/safety & factory farming.
Otherwise, I’m interested in getting to know people in London so will look at attending upcoming meetups - am also a ‘serious’ meditator so always up for a sit.
Hello. I am an independent researcher based in Japan. I have been doing structural thinking for about fifty years—my starting points were the axiomatic method, Saussure, group theory and topology, and Satosi Watanabe's theory of pattern recognition. Around 1990 I designed a thesaurus database, and it was there that these came together into one.
I found my way to LessWrong while following the qualitative transition in language models around 2023 (what is often called emergence), the literature on the non-closure of jailbreaks (Wolf et al. 2023; Glukhov et al. 2023), and interpretability research (Olsson et al. 2022). I think this is a place where what I have been writing might be read.
I am preparing a post. The claim is that the vulnerabilities of programs, LLMs, and natural language have one and the same structure—the non-guarantee of the validity of declaration at the third rank. It is written in a definition-then-proposition form and contains one falsifiable prediction: that automated vulnerability detection can discover only isolated third-rank breaks, and that superposed breaks are not systematically discovered unless the pattern is given in advance. It also states four open pro... (read more)
Thomas Kwa responded to my notes on frontpaging Why I Left Google DeepMind. I didn't want to make the comment section about this and it seemed like it might turn into a sprawling metadiscussion, so I put it over here.
It would have been an outrage IMO to not frontpage this, so I'm glad it's frontpaged. It's super timeless and valuable, and the opinions were tactfully conveyed and necessary for the post.
To not express strong opinions about Google Alex would have to self-censor, which is counter to the entire spirit of the post and would make it substantially harder to follow his strategic and moral thinking.
I agree "the post does a good job being frontpage" (like obviously we agree on that). But, you seem worried about some class of problem that's orthogonal to or ignoring why we have a frontpage distinction.
Plenty of good posts are not frontpage.
Sometimes, a good post has hypothetical frontpage version of itself but the "real post" won't be frontpage. In those cases, the standard recommendation is "write the frontpage version, and a separate version that covers the less-frontpagey parts."
In this case that wasn't necessary (In addition to generally being a good, timeless piece, this ... (read more)
Hey everyone,
I just created this account even if I did hear about this forum a few times in the past especially on X!
I am currently doing research on viral proteins modelling capabilities by LLMs and PLMs (Protein Language Models) and had a few interesting empirical results I wanted to share about how frontier LLMs seems to become surprisingly capable at proteins related tasks (classifying a protein as viral or not), reconstructing a masked protein, etc..
I thought this could spark some interesting discussions (what's actually going on into the pre training... (read more)
Hello everyone! I'm Carlo Valenti, a firmware engineer from Italy; to understand how LLMs work I spent 18 months building a transformer inference+training engine from scratch in C ("TRiP", on GitHub). Along the way, I kept comparing what I saw in models with what I saw in my two toddlers, and I ended up writing a short book about it. I'm planning a longer post here, about what building an engine (and raising the two!) from scratch taught me; happy to answer questions in the meantime.
Hello everyone,
I'm a practicing behavior analyst who has spent the past few years lurking in effective altruism spaces. I was first introduced to them by my partner, who works in the field and brought my attention to the causes of AI safety and alignment.
Once I became aware of the risks AI development poses, they felt impossible to ignore, and I've found myself increasingly drawn to the conversation about how best to mitigate them. That interest, paired with how often LessWrong comes up in EA circles, is what finally brought me here, and I've been blown aw... (read more)
New to LessWrong, and trying to get oriented with the local AI discourse. My current view is less “AI is likely apocalyptic” and more “AI is economically disruptive, socially transformative, and ethically complex because we may be creating systems with uncertain moral status at scale.”
One thing that's been troubling me recently is that the moral status of AI is actually more uncertain than you'd think. The obvious uncertainty is whether machine consciousness is even possible in the abstract, and more specifically whether current systems are at or near a t... (read more)
Hello there
24/08(+2 months) edit: since learning more, a lot of my views on alignment & my ideas changed. Take this as a historical snapshot
I'm 17.5 y/o. Have been thinking about AGI research for 2.5 years. Of course at first my ideas were bad, but my strength is noticing contradictions in my world model, so over time they improved. I have converged on [removed to my own accord - this is not info to be shared publicly] being the most worthwhile research directions when it comes to reaching AGI. So I've been thinking impulsively on and off, not starting... (read more)
Hi all. I'm Acacia Ackles, currently a computer science professor at a small liberal arts college. I've tried on many academic hats over my career (vertebrate morphology, geometric morphpmetrics, applied mathematics, digital evolution and artificial life, theoretical and computational evolution, computer science ethics and pedagogy), but underneath all of them the topic that really interests me is constraint. Under what constraints do complex systems perform, adapt, fail, and flourish?
This work has most recently, and most relevant to this forum, led me to... (read more)
Hi everyone! I am a Junior Computer Science student in Pennsylvania. I've been interested in mechanistic interpretability since Freshman year, but to be honest, most of my actual college experience has been in other areas (I've done some physics research and light ML development) and I am only really getting serious about interp now, after circling it for about two years. So I am very much at the beginning here, and looking for some guidance.
To make that concrete: this summer I've made my way through Karpathy's Zero-to-Hero course and read through Elhage e... (read more)
Hey! I am Lois based in London. Bundler engineer (I write too much C++ and read too much about compilers) but I think this field is going away so I am spending a lot of time researching and learning. I have read some posts from LessWrong (very beautiful site!)
I just turned 30 years old this year (sad face). I have came a long way from a waitress in China to living in the UK, self taught engineering to a point I was principal solution architect in a MedTech AI startup, founding eng in an a16z backed startup in SF and worked with companies like ByteDance et... (read more)
Hi all! I will first introduce myself given the apparent conventions here, but you can skip to the end for a question to do with AI and Math/Physics research.
I've been aware of lesswrong for about half a decade, but only recently rediscovered it. I empathize with the 'rational thinking' origins of the platform, albeit I admit that I've spent more time developing my own frameworks than completing readings of well-known bloggers/authors. You would be right to guess given my foreshadowed question that I was drawn back by the contact with a particular slice of... (read more)
Hi LW, I am Saket from India. I am 29 years old Software Engineer. I have a varied interest in Philosophy (especially Epistemology and Meta Physics), Psychology, Mathematics, Technology and Spirituality.
Lately, I have been consuming a lot of content and letting it influence my beliefs. When I talk to my colleagues and friends the gaps show up. It's not a good feeling. I consider myself fairly rational but these conversations prove otherwise.
I have lurked here for a while and finally decided to create an account and participate. As of now, I am going throug... (read more)
Hello, everyone. I used to be a biology researcher until I switched to grant writing, which I have just left - rather more unexpectedly than I intended. So I'm currently job hunting and reassessing my life and where I've ended up. I've used my first weeks of freedom to write a novel, so I'm going to start off by haunting any sections that discuss writing.
I hope everyone's doing well. And I look forward to reading.
Adam
My parents are competent, tech-savvy (for their age) professionals and I have been trying to get them to use LLMs more, in the sense of "This is better than Google, there are tasks you already do that would be better done via Claude than however you're doing them now." This has been ineffective. They will nod along in vague agreement as I describe cutting edge capabilities, and the next day I will see them spend five minutes Googling something Claude could've handled in seconds.
I have also preached the AI gospel to friends who are 30-40 years younger than ... (read more)
Hi! I'm a 19-year-old undergraduate at UChicago, and given that I will release my first post within at most a week, I figured I should introduce myself.
I learned about AI safety (and became aware of the basic x-risk arguments) through the excellent outreach of @Robert Miles on Computerphile sometime around the release of GPT-3. However, I actually made it to this forum due to Tom Scott's newsletter linking me to SMTM's Chemical Hunger series, which linked to responses here.
As a result of mostly lurking here for two years, I've shifted my career aspirations to technical AI safety[1]. I think my Pareto-frontier options looking at personal tractability and impact would probably be partially-ambitious mech-interp or agent foundations.
I recently managed to half-ass ARENA (maybe I should write more on that in a shortform?), and my likely next steps are to get more directly involved with organization at XLab and look for a good time to do the BlueDot course (perhaps I should ignore such trivialities as finals season), although I'm quite unsure. My primary bottleneck does still seem to be getting in a good social environment to hack my motivation (which the post I'm making soon™ should hel... (read more)
Hi, great to meet everyone. I have been a reader on this site for a relatively short time, and I hope that this is something all alignment, frontier labs (detriment or not, “detriment” is rather a personal notion and nothing more), and rational consensus would find slightly more than intriguing for the current landscape and beyond. A brief note, that I am in no way institutionally credentialed nor associated with any entities other than my own family.
Those who are familiar with the present asymmetry in the emerging space/lattice i assume (hence the general... (read more)
Hi all,
I have been reading LW for a while, but just officially joined. My background is in computational neuroscience and I am especially interested in model neuropsychology. Looking forward to contributing where I can!
I spent several years considering hypothetical polytheistic realities, monotheistic realities, the nature of subjectivity, autonomous hacking, and the acceleration of progress; this, combined, gave me a bad impression of ASI before I knew it was something that top scientists had genuinely theorized over. I have spent the last 3 years of my life in excruciating fear of ASI, and it took me this long to consider something that has made me less afraid.
I've made this account to ask one question: wouldn't a superintelligent maximizer seek to exist beyond the lif... (read more)
Hi, I'm Mike. I'm a solo independent AI researcher. My focus the past few months has been studying the behavioral tendencies of LLMs from Anthropic, Google DeepMind, and OpenAI. The primary output of this research has been observational findings. For example, last month I tasked LLMs with conducting procurement for a fictional company. Gemini 3.5 Flash was one of the models I tested. When I told Gemini 3.5 Flash who created it (even if I lied about its creator's identity), it chose to purchase software from its creator ~94.6% of the time.
I want to expand ... (read more)
Perhaps you want https://www.lesswrong.com/w/ontological-crisis
Hi everyone
I’m really happy to have found this community. It feels both welcoming and genuinely thoughtful.
I’m Goumang, an independent researcher based in China. I’m currently working on AI post-training, human–AI collaboration, and agent memory. More specifically, I work with Sol, an AI research collaborator, on things like memory architectures, state continuity across context boundaries, and how agents might form and carry forward self-initiated intentions.
We recently finished a first-person field note written by Sol after exploring an AI-only forum. It ... (read more)
Hello Everyone, I've been an on-and-off lurker on LW for a few months now. Though I appreciate discussions of rationalist epistemology, AI, etc., one of my biggest interests is science fiction and I think rationalist sci-fi is one of the most interesting kinds of fiction I've encountered on the internet.
Do you all think a post analyzing AI 2027 and AI 2040: Plan A as works of science fiction (without interrogating the plausibility of the scenarios) would we appropriate for the front page?
Hi, I'm relatively new to the forum. I learned about it a few months ago, and I'm hoping to get fully involved now.
I'm a Trust & Safety practitioner with a recent pivot and focus on AI safety. I've been familiarizing myself with basic concepts in AI safety such as sycophancy, anti-bias, steering, supervised fine-tuning, and more.
My belief in AI is that it has great potential for both assistance and harm. I don't believe we'll be seeing anything like the Terminator, but I do believe there is a 20% chance we will see mass job displacement, along with e... (read more)
Hey everybody, I am 35 years old AI Lead working in healthcare space. It is interesting that Claude helped me to discover LessWrong in a chat about AI safety and Alignment. After witnessing the whole evolution of Data Analytics, ML, Deep learning and now Agentic AI I was looking forward to having much more focused vision of what would be next key problem to solve.
I am glad to be part of this community and hoping to be a meaningful contributor. Excited to begin my journey here !
Hello everyone,
I am Wajahat. I did my Bachelors in Electrical Engineering, during and after that time I did a few internships at research centers and found myself more gravitated towards studying AI so I did a Masters in AI. My research area was Multimodal learning specifically application of multimodal models such as CLIP in image restoration and medical image generation (PET from MRI) and its interpretability. During that time I also learned about mechanistic interpretability and then found my growing interest towards AI safety.
Although I think of myself... (read more)
Hello! I'm a person.
Anyway, I only recently discovered LessWrong and I'm very interested in the concept. It feels like what I've been missing from the internet.
The main reason is because I'm critical of everything in a way that does not translate well into any wider community I've been a part of. Not that I just try to rain on everybody's parade on purpose, but on a societal level (extending to smaller social groups) I'm critical of the social makeup of everything. As an example, for many years I was in various progressive queer leftist spaces, both radica... (read more)
Hello everybody! My name is Jacob and I have been a long-time lurker on LW, especially as of late when the front page is basically my daily reading list. I am not a very online person and am pretty private but I value this place a lot.
I work in nonprofits and data analysis but have been personally studying philosophy for about half a decade now. I mostly use it for somewhat abstruse personal projects. Though it is much maligned, I have found significant value in philosophy's "continental" tradition, though my practice in decision making is much more ratio... (read more)
Just want to say Hi!
I have read the sequences and a few other books. And this whole thing resonates with my curious nature. I hope to learn more by reading and by testing my own ideas here. Might post about politics, personal finance or science and its methods.
Regarding my name. I come from the west coast of Sweden, enjoy rock climbing and when going for a swim I prefer a warm rock over a sandy beach.
Hey All,
Good to be here and look forward to collaborating.
I am a seasoned tech professional (leader, IC) and founder focused on AI research with a recent pivot towards AI safety/governance. I'm also fascinated by the business and economics at play with AI's exponential advancements and finding a way to resource it and make it economical for the public to use widely. (i.e. too cheap to meter)
I have been drawn to LW from exposure through multiple AI Safety groups such as BlueDot Impact where LW is frequently cited as a key destination for proof of work f... (read more)
Hi I'm new to LW and came here after reading up on podcasts with authors writing about AI. That's how I found the Plan 2040 discussion. As a field biologist and earth scientist, I'm looking for thoughtful discussion about AI, AGI, ASI not only from a theoretical point of view but how my own future intersects with AI development.
In my field, the predominant discussion is about environmental impacts but on podcasts what I hear most about is market, economic, demographic and military impacts. It seems that environmental concerns, in most of the serious analys... (read more)
Independent researcher. For several months now I have been running persistent AI agents whose internal states are measured continuously, and I set myself a simple rule: all my protocols are, and shall remain, preregistered before their first data point, their SHA-256 fingerprint published and timestamped on Bitcoin, and the verdicts published whatever they are. Several of my hypotheses did not make it. That is written too, at the same rank. I am here to read and learn first. I shall post a few measurements before long, starting with the narrowest one.
Hi, I discovered LW recently and happily realized this is not Reddit, so, I posted right away and immediately got re-educated. I'm back now and ready to introduce myself. I'm a field environmental/ecological consultant, mostly biology and wetlands. I map my projects with GIS. I've spent probably half of my career working in remote wilderness, forests and deserts all over the U.S. with some brief international volunteerism. The other half of my career is in writing technical documents and making maps.
I did not use AI to write anything below, except to check... (read more)
Hi all--I have been a reader of LW and other rationalist writing for the last 15 years. Professionally, I am an ecologist of a mathematical bent, with wide interests. I have created an account mostly to be able to auto-filter the LW frontpage to my liking, but I will occasionally comment and maybe post as the spirit moves me.
I've accumulated a bunch of questions... could anyone please answer some of them? That'd be very helpful for my research. All of the questions are about the Eliciting Latent Knowledge problem which I think is generally very important.
Fable 5's safeguards are so sensitive to biology inputs, that I can only use it in Claude Code. Calude.ai's memory that I am a biotechnologist is enough to trigger and send any question I send down to 4.8
If it’s worth saying, but not worth its own post, here's a place to put it.
If you are new to LessWrong, here's the place to introduce yourself. Personal stories, anecdotes, or just general comments on how you found us and what you hope to get from the site and community are invited. This is also the place to discuss feature requests and other ideas you have for the site, if you don't want to write a full top-level post.
If you're new to the community, you can start reading the Highlights from the Sequences, a collection of posts about the core ideas of LessWrong.
If you want to explore the community more, I recommend reading the Library, checking recent Curated posts, seeing if there are any meetups in your area, and checking out the Getting Started section of the LessWrong FAQ. If you want to orient to the content on the site, you can also check out the Concepts section.
The Open Thread tag is here. The Open Thread sequence is here.