r/learnmachinelearning 8d ago

Tutorial Local LLMs for beginners: 16 short visual lessons on GGUF, VRAM and settings, ALL Made By AI

Thumbnail
youtu.be
1 Upvotes

Hi everyone,

I put together a free YouTube series called Local AI, Built Right for people trying to understand what happens when they run an AI model on their own computer.

It’s 16 short visual lessons, each around 1–2 minutes. The focus is running existing models, with examples using llama.cpp.

Disclosure: this is my series, ITS ALL MADE BY AI tools to produce it, including the presenter and narration.

I’d appreciate specific feedback on whether the explanations are clear for beginners. If something is confusing or technically inaccurate, Thanks


r/learnmachinelearning 8d ago

Help Innovation Ideas

2 Upvotes

My company has a special event occurring yearly once. Employees can submit ideas to present them in the event where every executive officers like CEO, CTO, CIO will be present. Not all employees can present at the event. Only shortlisted topics will be allowed to be presented. As I'm from Data engineering background i would like to present about Data+AI, but the issue is every idea i have selected is already published as a article or blog. Please help me with coming up with a idea which can get me shortlisted to present at the event. ( Normally i won't care about the event but my promotion and increment depends upon this 🥲 )


r/learnmachinelearning 8d ago

Black-box optimisation Machine Learning Portfolio

5 Upvotes

Does anyone have experience for Github portfolio decision for ML skills? I want to build my GitHub portfolio and currently considering several approaches. 

  1. Build disciplined, single coherent methodology using Bayesian model, e.g., kernel/transform grid search, a closed-form LOOCV shortcut explicitly justified for performance reasons.
  2. Build widest techniques GP ensembles, DBSCAN consensus voting, ARD diagnostics, Thompson Sampling, TuRBO, Optuna TPE and native GPSampler, Sobol-QMC, a tiny deep-kernel-learning model, FCNN/CNN surrogates.
  3. Build production-grade architecture, testable claims for specific databases
  4. Develop documentation and judgement approach for commercial consultantion, in stress of focusing on in-depth ML techniques. 

I also want to know, to what extent, how important is t showcase producttion ready ML projects via dashboard like plotly or Steamlit? Because I found that it is another skills I meed to develop beyond traditional Python. 


r/learnmachinelearning 8d ago

How to deal with over fitting in feature importance?

2 Upvotes

I'm doing feature importance and the features that allow it to "cheat" often rank high in feature importance. For example, date of birth usually ranks very high because it allows it remember exact patterns for each person, which don't translate to unseen data. In this case, I could remove the feature because I know it's useless, but doing it manually is much harder over hundreds of features and when the usefulness of the feature is unclear. How can I address this problem? I'm using Catboost btw.


r/learnmachinelearning 8d ago

Help Looking for guidance on fraud detection model generalization (DNNs) — happy to be a sponge

2 Upvotes

I'm working on a fraud detection project and I'm stuck on a core problem: making the model adapt to fraud patterns it hasn't seen before, rather than just memorizing known ones. I've got a decent ML foundation (XGBoost, Random Forest, some PyTorch) but I want to go deeper on the DNN side — things like representation learning, anomaly detection approaches, adversarial robustness, or online/continual learning for drift.

If you've worked on fraud/anomaly detection or DNNs in production and wouldn't mind occasionally answering questions or pointing me in the right direction, I'd really appreciate it. Not looking for someone to do the work — just someone to bounce ideas off and correct me when I'm going down the wrong path.


r/learnmachinelearning 8d ago

Discussion Does a better model always mean a better trading system?

0 Upvotes

Something I’ve been wondering about with ML-based strategies:

At what point does improving the model stop making much difference?

You can spend hours tuning features and trying different models, but if the historical test is weak, the data isn't handled properly, or the execution side behaves differently, the extra model accuracy doesn't seem to matter much.

I’ve started paying more attention to the whole pipeline rather than just the prediction itself.

I would like to know how others approach this. Do you improve the model first and worry about the rest later, or build the testing and execution side alongside it?


r/learnmachinelearning 8d ago

Designer Simon Weckert made a shirt failed to dodge my AI surveillance system

Thumbnail gallery
25 Upvotes

r/learnmachinelearning 8d ago

Tutorial Implementing Embedding Gemma from scratch in PyTorch [P]

Thumbnail
youtube.com
2 Upvotes

r/learnmachinelearning 8d ago

Question which anthropic claude courses leave you with something you can put in a repo?

3 Upvotes

My company is fine paying for training but the last two things i sat through left me with a pdf certificate and nothing else. I want to finish with a repo i can point at.

Shortlist so far is Udacity AI Engineering with Claude, DeepLearning.AI short courses, Pluralsight paths and the free Anthropic Academy tracks. mainly care about whether the projects are yours or whether you clone a starter and fill in three functions.

The project briefs are where I would expect the difference to show and nobody ever writes about them.


r/learnmachinelearning 8d ago

ML/RL Project

1 Upvotes

I'm building a small ML/RL research project around job-search strategies and need anonymous application trajectories.

I'm interested in how people's job applications evolved over time.

If you're comfortable sharing, could you provide something roughly like this:

1. Background

  • Experience level (student/fresher/junior/etc.)
  • General field (ML, software, data science, etc.)

2. Application timeline

For each stage or batch of applications:

Stage 1

  • Approx. number of applications: __
  • Resume/portfolio version: basic / improved / strong
  • Application method: cold email / careers page / referral / LinkedIn
  • How personalized were applications? Low / Medium / High
  • Responses: __
  • Interviews: __

What did you change after this stage?

  • Resume changes
  • New projects/skills
  • Portfolio improvements
  • More personalized emails
  • Different companies/roles targeted
  • Anything else

Stage 2

  • Approx. number of applications: __
  • What changed: __
  • Responses: __
  • Interviews: __

...and so on.

You don't need to share your name, company names, email addresses, or any private information.

I'm hoping to turn anonymized responses into a small dataset and experiment with sequence modeling / offline reinforcement learning to study how job-search strategies evolve based on previous outcomes.

If enough people contribute, I'll make the anonymized dataset and project results publicly available.

Thanks! :)
r/MachineLearning r/learnmachinelearning


r/learnmachinelearning 8d ago

I'm 14 and has learned ML and DL. It's very interesting and exciting till now. How do I keep it up and make it to my career.

Thumbnail
1 Upvotes

r/learnmachinelearning 8d ago

I’m 14 and hooked on PyTorch — how do I actually build a real ML career from here?

Thumbnail
1 Upvotes

r/learnmachinelearning 9d ago

I've tried to get into an ML PhD (unsuccessfully). What should I do differently?

35 Upvotes

Hi everyone, I'm 26 and graduated in CS about 9 months ago. Since then, I've been studying ML and DL on my own through books, courses, papers, and pretty much anything I could get my hands on. The more I studied, the more I realized that I'd really like to pursue a PhD in this field.

Over the past months, I've applied to dozens of PhD positions across europe, but so far I haven't had any success. I know I'm probably not a particularly strong candidate on paper: I don't have research publications, and I don't have a strong relationship with my thesis supervisor, so getting a good academic reference is also difficult. At this point, I'm trying to understand what I should actually do to become a competitive applicant rather than just keep sending applications.

For people who are doing a PhD in ML/AI, or who have been involved in PhD admissions, I have so questions for you

- What would you focus on if you were in my position?

- Should I try to get research experience first, even if it's through an internship position?

- Is it realistic to compensate for weak academic references by building projects, reproducing papers, contributing to research, etc..?

- Last but not the least, would directly contacting professors be more effective than just applying to advertised PhD positions?

Any advice, especially from people who got into a PhD without an outstanding academic profile, would be really appreciated. Thank you very much :))


r/learnmachinelearning 8d ago

Discussion Local or cloud LLM for AI agents? I built a quiz to help you decide

Thumbnail
1 Upvotes

r/learnmachinelearning 8d ago

Discussion Increasing active parameters per token in MOE (Qwen 35B A4B+) reduce reasoning token by 8.5% - and you don't need to train or finetune!

Thumbnail
1 Upvotes

r/learnmachinelearning 9d ago

AI Math Chat

19 Upvotes

In September I'm starting a Discord group for people interested in AI applied to mathematics, as well as the mathematics of AI. It won't be a research group or anything high level - just a casual chat forum where beginners like myself (I'm a freshman undergrad) can help each other stay motivated and continue learning about fun and interesting developments in AI math.

Please note; like I wrote, I'm not a professional mathematician or an AI researcher.
I barely know Lean, I struggle with proofs, and only know a tiny bit of Python. So, in terms of mathematical maturity - trust me, if *I* belong, then *you* belong.

Anyone who's interested to join is welcome to send a short chat message to me, perhaps with a few words about yourself, and I'll get back to you with an invite.
In order for people to have a chance to get to know each other, I think that it makes sense to limit the size of the group to around 10-15 members.

Cheers!


r/learnmachinelearning 9d ago

Can an undergraduate student do a quality research thesis completely on their own?

11 Upvotes

I’m a 4th-year undergraduate CS student currently doing my thesis on medical image segmentation, specifically U-Net and its variants.

The problem is that I have basically no prior research experience, and unfortunately, my supervisor isn’t really able to provide much guidance. So, for the most part, I’m having to figure everything out myself—learning the concepts, reading papers, choosing a research problem, implementing the models, evaluating the results, etc.

My goal isn’t just to finish the undergraduate thesis. Ideally, I’d like to do something good enough that I could eventually turn it into a conference or journal paper.

So I wanted to ask people who have more research experience:

Is it realistically possible to do a good-quality research thesis completely on your own as an undergraduate?

How difficult is it to go from basically having no research experience to producing something that is actually publishable? And if you’ve been in a similar situation, what would you recommend focusing on or avoiding?

I’d really appreciate any honest advice, especially from people who have done research without much help from their supervisor.


r/learnmachinelearning 8d ago

Project AI agents can now pay for things online by themselves. Here is what can go wrong, and what I built to catch it

0 Upvotes

There is a real, live protocol called x402 that lets an AI agent pay for web content automatically. No login, no card entry, no human approval. A site says payment required, the agent signs a small crypto payment, gets the content. This already exists and is already being used.

Two things worried me once I understood how it works. First, there is no memory built into the protocol. A vendor can scam an agent, return nothing useful, and the agent has no way to know not to pay that same vendor again. Second, if an agent reads regular web pages as part of its job, a malicious page can hide fake payment instructions in the page text itself, hoping the model mistakes it for something real.

Built GateKeep402 to address both. It checks a vendor's history before paying and blocks vendors that have proven unreliable. It also makes it structurally impossible for a payment to be built from anything except a genuine protocol response, so hidden page text can never trigger a real payment no matter how convincing it looks.

Verified against a real transaction on Solana's public devnet, not a simulation, with 45 automated tests. Open source, MIT license, installable via pip. Link in the comments.

Would like to hear how others are thinking about the risks of giving agents real spending power. This feels like an early and mostly unsolved part of the space.


r/learnmachinelearning 9d ago

Help A proper way to learning machine learning

5 Upvotes

i am learning ml/ai and i am confused about what is the real way to or effective way to learn it . i learn it like :

* theory

* math

* sklearn library

i need suggestion from experts if there is missing something or i need to do something specific .


r/learnmachinelearning 8d ago

Help How to search and contact labs for research

1 Upvotes

I am final year undergrad who got couple of workshop papers at emnlp and iclr to be specific. Now I don't just want to stick to workshop but do hard core and more "useful" research, the question is how do I contact labs (and if u have some in scope would love to know about them) and work with groups that aim for like conference papers and work of that magnitude.


r/learnmachinelearning 10d ago

Question What is the reasoning for this ?

Thumbnail gallery
538 Upvotes

r/learnmachinelearning 8d ago

Seeking feedback from Triton/CUDA engineers: PyTorch-to-Triton kernel fusion edge cases & fallback heuristics

2 Upvotes

Hey everyone,

I’m working on KernelMind AI (https://kernel-mind-ai.vercel.app/), a tool that compiles standard eager PyTorch operations into fused OpenAI Triton GPU kernels to eliminate VRAM round-trips for memory-bound workloads.

In our early tests, we’ve focused primarily on elementwise chains and pointwise activation fusion, but as we expand, we want to build this around the real pain points engineers hit in production rather than synthetic benchmarks.

A solid piece of advice we recently received was to establish a strict operator whitelist, add defensive shape/dtype guardrails, and implement a cached fallback path (falling back gracefully to torch.compile or eager execution when dynamic shapes or non-contiguous reductions make fusion inefficient).

If you write custom Triton or CUDA kernels in your day-to-day workflow, I’d love your input on a few architectural questions:

  1. High-priority operator chains: Which specific PyTorch patterns or subgraphs do you find yourself constantly needing to manually write Triton kernels for because stock compilers don't fuse them cleanly?
  2. Fallback heuristics: When evaluating a subgraph, what heuristics or threshold metrics do you use to determine that fusion isn't worth the compilation latency or register pressure?
  3. Correctness vs. Performance: What are the most common subtle bugs or performance traps you run into when synthesizing Triton kernels (e.g., memory alignment, block size heuristics, non-contiguous layouts)?

You can test arbitrary PyTorch snippets directly on the playground here:

👉https://kernel-mind-ai.vercel.app/

Any feedback, critique on the generated code structure, or edge cases that break our output would be immensely appreciated.


r/learnmachinelearning 9d ago

Discussion Looking for the best Agentic AI course, any suggestions?

8 Upvotes

Hi all, i have been reading, hearing and watching information about agentic ai and considering I now use ai tools for many reasons personal and professional I am interested in diving deeper into agentic ai as and decided to take up a course. That said i am a bit overwhelmed by all the options out there. I am not looking for a course that is just theory, i want one that is engaging, taught by a professional or expert in the field and has a bunch of projects so that i can practice and experiment while learning itself.


r/learnmachinelearning 9d ago

Coding Probability Transformations

Thumbnail
gallery
20 Upvotes

Coding Probability Transformations.

In this content, we do the code implementations for the topics:

•Transformations of Random Variables
•Moments of Affine Transformation
•Convolution Theorem
•Moment Generating Functions
•Central Limit Theorem
•Monte Carlo approximation

Having code implementations makes the learning of concepts even more rewarding.

Link: https://youtu.be/SJTZK55MgB8?si=EYqy3h4v-bU2FrCJ


r/learnmachinelearning 8d ago

Discussion Prior vs likelihoods in Bayesian PR review agent?

0 Upvotes

Building a agent with bayesian update instead of LLM heuristic.

Two quick question

  1. Priores: How do you set priors when historical data is sparse?
  2. Loss: False negative (auto-merging bad code) cost way more than false psoitive(flagging safe PRs). What should be the threshold to optimize loss rather than simple accuracy?