Hacker News
new
past
comments
ask
show
jobs
points
by
salamo
1 days ago
|
comments
by
scoriiu
1 days ago
|
[-]
did you skip simd just because the model's tiny? naive conv perf is honestly the only reason i haven't done exactly this for the cnn
reply
by
salamo
1 days ago
|
parent
|
[-]
Yeah, the model is small enough that inference is already basically instant for my usecase (only 6 transformer layers for the blog search).
reply