We thought sigmoid activation function is best for more than 50 years. Just recently if you draw a straight diagonal line in the plot, we found that it significantly outperforms the other and it is called ReLU. Considering how long we have been doing wrong for half decades on a simple problem, how wrong is the current state of ML?
We thought sigmoid is best
Kyle(120.18)
2018-03-12 17:07
추천 0
댓글 10
다른 게시글
-
visual basic 안깔고 C나 C++하는 방법 있나 [4]ㅋㅋ(218.151) | 18.03.12추천 0
-
형들 ㅠㅠ zone 이랑 dight 쉽게설명좀해주라익명(110.70) | 18.03.12추천 0
-
제명유 come back! 슈퍼스타 제명유가 나타났다위키세계어(angel11511) | 18.03.12추천 0
-
포폴을 제대루 한번 만들어봐야게다 [1]Lunatic(39.7) | 18.03.12추천 1
-
책 추천좀 해줘엉 asm [2]하앙(175.223) | 18.03.12추천 0
-
참고로.. 1교시 1번 문제에 대한 참고용 답지. [20]☎2.80™(roidz) | 18.03.12추천 0
-
c# 책 추천좀 [4]minwoo2815(minwoo2815) | 18.03.12추천 0
-
그새끼 언제 뒤지냐익명(211.109) | 18.03.12추천 0
-
걍 C만 제대로해도 함슬람따위가 하는거 더 고성능으로 구현가능 [5]익명(110.70) | 18.03.12추천 1
-
ㅋㅋ 플밍에 자격증이 있다는게 더 웃기지. 그래서 전화기가 틀딱인거야 [7]위키세계어(angel11511) | 18.03.12추천 2
안녕~
안녕
This is false. ReLU has its pros and cons.
Uh oh, we have 낄낄 clone nay-sayer
Well, saying that doesn't make this true. sigh. can you please google before you make this kind of dumb statement?
It's a statement from Geoffrey Hinton you moron
I can hardly believe that's true. Keep saying that to yourself. If you have ever googled on that particular subject, you'd know that's an embarrassing statement.
Sounds like an expert. If you double the number of layers in an NN, how much more time do you expect the backprop to take?
Answer format: times <approximate number>
btw I'm talking about a Deep NN