this post was submitted on 15 Jun 2023

3 points (100.0% liked)

Learn Machine Learning

530 readers

1 users here now

Welcome! This is a place for people to learn more about machine learning techniques, discuss applications and ask questions.

Example questions:

"Should I use a deep neural network for my audio classification task?"
"I'm working with a small dataset, what can I do to make my model generalize well?"
"Is there a library available that implements function X in language Y?"
"I want to learn more about the math behind machine learning technique A, where should I start?"

Please do:

Be kind to new people
Post guides and tutorials that you find helpful
Link to open/free sources instead of paywalled when possible

Please don't:

Post news articles / memes (there are other machine learning/AI communities for this)

Other communities in this area:

Similar subreddits: r/MLquestions, r/askmachinelearning, r/learnmachinelearning

founded 2 years ago

MODERATORS

[email protected]

[email protected]

3

[Repost Q] Should I scale values before using them to train autoencoder? (sh.itjust.works)

submitted 2 years ago* (last edited 2 years ago) by [email protected] to c/[email protected]

9 comments fedilink hide all child comments

Not OP. This question is being reposted to preserve technical content removed from elsewhere. Feel free to add your own answers/discussion.

Original question:

I have a dataset that contains vectors of shape 1xN where N is the number of features. For each value, there is a float between -4 and 5. For my project I need to make an autoencoder, however, activation functions like ReLU or tanh will either only allow positive values through the layers or within -1 and 1. My concern is that upon decoding from the latent space the data will not be represented in the same way, I will either get vectors with positive values only or constrained negative values while I want it to be close to the original.

Should I apply some kind of transformation like adding a positive constant value, exp() or raise data to power 2, train VAE, and then if I want original representation I just log() or log2() the output? Or am I missing some configuration with activation functions that can give me an output similar to the original input?

top 8 comments

sorted by: hot top controversial new old

[–] [email protected] 3 points 2 years ago (1 children)

yes, simply rescale proportionally.

[–] [email protected] 3 points 2 years ago (1 children)

Wow, you're quick! Beat me to reposting the old answer

[–] [email protected] 3 points 2 years ago (1 children)

ty for your work! hope you don't mind me reposting these to my instance. happy to go both ways too!

[–] [email protected] 3 points 2 years ago (1 children)

Ya, go ahead. Atm I'm just pulling old answers that I've deleted from reddit or cross validated. I'll probably do some underappreciated/hard to find answers as well

[–] [email protected] 3 points 2 years ago* (last edited 2 years ago) (1 children)

this is the real work we need right now, one of the biggest complaints is the searchability of usable content on reddit (lemmy does not have much), we know the answer to that.

I have a list of AI/ML folks I am following an posting fresh research, we are using bots to summarize the work as well. Feel free to repost with abandon as well.

[–] [email protected] 3 points 2 years ago (1 children)

I think more of the traffic comes from google search results (that's what people say they use) and there's not much that can be changed there.

Imo there is too much noise in the ML research publishing atm, certainly with deep learning (throw an NN at something and it seems like you can get anything published nowadays) and the papers aren't very helpful (they rarely have enough info to be reproducible) but I digress. Plus paywalls can be a problem. If you want the more solid stuff for deep learning, I would suggest something like: paperswithcode.com/ But I spend most of my time sorting through papers so not an unbiased opinion at all

[–] [email protected] 3 points 2 years ago

Feels, integrated AI started as a shared google doc of links for some projects i was on, realized I had figured out how to wadge through papers and get usable stuff still (used to do similar for industrial robotics years ago).

My hope is as our domains age we can play SEO games and get our instances in the results. Make them look at us on every page!

[–] [email protected] 2 points 2 years ago

Original answer:

Scaling the dataset before passing it to the autoencoder is usually how I do it, you don't need to rescale after if you are only using the encoder portion (for example for dimensionality reduction). If you don't do it linearly (aka (x-min(x))/(max(x)-min(x) ) and use exp or log to do it then be mindful that it would likely have an impact with respect to loss/optimization behaviour.

Make sure to take the max and min values from the training data then apply it to the testing (in the case of values out of bounds, set them to the boundary value but this shouldn't have a big impact if your training dataset is large enough with enough variance).

load more comments