402 karma · joined March 19, 2007
Took months to get it all back.
If you're looking for a higher-level approach a lot of the community likes sci-kit which has lots of nice defaults for popular models.
Since you're familiar w/ the youtube paper, I've been wondering this question: How do they get vectors out of the softmax?
I can train on multi-GB datasets w/ only lightFM and multiple CPUs.
Another interesting package is called Implicit. This package, although not as complete as LightFM when it comes to algorithms or APIs, really shines when it comes down to optimizations. Including native Cuda kernels for BPR and ALS it also has an important speedup called the Conjugate Gradient Method which makes it faster than spark in some benchmarks.
But usually, now-a-days my work requires more customized hybrid models of which I usually start w/ a base BPR implementation I have in Keras.
Look at a talk from a Dutch group called Blendle at RecSys conf where they talk about these problems.
What are these patches and what was your intuition on starting w/ it?
You also want to deploy this on a standard machine w/ or w/o a GPU? That can have implications on inference speed. But there are ways to optimise for that. The absence of a network connection is not a problem.
Also, do you have training data? If not we can probably leverage a pre-trained model for this use case where, in this case we would only need a handful of training examples.
To answer some of your questions:
Where should I look to find a quality freelancer? Not sure.
What formats should I specify for the deliverables? Depends on what tools the developer uses, for example, I would deliver files in TensorFlow export format.
What timeline should I expect? Using a pre-trained model, I could expect to deliver it in about 40 man hours.
What pricing should I expect? Probably about $6000, but that depends on the contractor.
What would a good developer expect me to provide in terms and training data? Using a pre-trained model you would only need a handful of training images. But, of course, the more the better.
What API parameters would they expect me to specify? I can't think of anything.
I have experience building and deploying these vision models. Feel free to reach out to me for more info. It's my username @ gmail.com
As long as you diligently use your calendar and keep notes for each company it's pretty easy to manage this.