HNHacker News
TopNewBestAskShowJobs

vatican_banker

131 karma · joined June 23, 2016

submissionscomments
vatican_banker··on Program-of-Thought Prompting Outperforms Chain-of-Thought by 15% (2022)
DSPy implemented program of thought since a long time ago and it works great to solve user queries with code.

What is great is that you can define DSPy signature of the type “question, data -> answer” where “data” is a pandas dataframe, then DSPy prompts the llm to answer the question using the data and python code. Extremely powerful.

vatican_banker··on Ask HN: What are some comfy/stress-free jobs a SWE can do? (LCOL country)
What hobbies are you passioned about? You hobbies is usually a good place to start looking.
vatican_banker··on Ask HN: Want to leave my job with nothing lined up
Do it!

I was in your situation in 2023 and decided to quit without a job lined up. I rested for 2 months and then, out of nowhere, I had energy and motivation to start exercising again and start working on my mental health (with the help of a therapist/professional coach). I would recommend to have one single expectation if you decide to quit: rest.

All in all, quitting the toxic workplace and taking time off was the best thing I’ve ever done.

It took me 7 months to get an offer, 9 to actually start working.

Be prepared to be out of work for ~1 year.

vatican_banker··on Our $100M Series B
Banks care _a lot_ where the data is store, hence most banks are inclined to keep customer data on-prem.

Because data must be on-prem, banks are stuck in legacy infra paradigms. The whole org suffers, innovation is stiffled, yada yada…

An on-prem cloud product (hardware+software) is a game changer for these companies, IMO.

My question to oxide: how easy is to integrate external hardware into the cloud? For example: bunch of GPUs or a bunch of next-gen hardware like SambaNova.

vatican_banker··on RouteLLM: A framework for serving and evaluating LLM routers
The tool currently allows only one set of strong and weak models.

I’d be really good to allow more than two models and change dynamically based on multiple constraints like latency, reasoning complexity, costs, etc.

vatican_banker··on RouteLLM: A framework for serving and evaluating LLM routers
> Trained routers are provided out of the box, which we have shown to reduce costs by up to 85%

The answer is here. This is a cost-saving tool.

All companies and their moms want to be in the GenAI game but have strict budgets. Tools like this help to keep GenAI projects within budget.

vatican_banker··on Show HN: Struct – A Feed-Centric Chat Platform
Countless tools have promised to highlight "what's worth my attention". None of them have worked for me or my team. What's different on Struct?

A few questions I jotted down while watching the video on Struct's landing page:

1. the concept of channels seems to be important on Struct as channels are the starting point of threads/feeds. Could you clarify the concept of channels on Struct? Is it just a concept to group users? Can you also chat on channels?

2. Conceptually how do you handle the fact that only the threads on the realtime feed are visible to the user? Maybe there's a low-signal high-activity thread that takes space and hides the high-signal low-activity thread which results in users missing important information or reminders.

3. Tags are crucial for filtering threads, is there a way to "police" the tags? Using tags usually grow into a mess of similar-but-not-the-same collection of text. Think of JIRA tags.

4. How to handle threads created independently by different users but discussing the same topic?

5. Not a question, but I'd be interested in knowing more about private conversations between two parties. It's mentioned only briefly in the video.

Hopefully these questions don't come out as overly critical. The tool definitely has potential.

vatican_banker··on BLIS: A BLAS-like framework for basic linear algebra routines
Embarrassingly parallel means "singe task multiple data" or "multiple task single data". Canonical example: monte carlo simulations.

With that out of the way, basic linear algebra operations that require sophisticated algorithms and are not "embarrassingly parallel": matrix multiplication, matrix inversion, matrix decomposition (SVM, QR, etc). Some of these algos fall into BLAS and others in LAPACK.

vatican_banker··on BLIS: A BLAS-like framework for basic linear algebra routines
> numerical algebra work [...] is mostly embarrassingly parallel

It's the exact opposite, most numerical linear algebra is _not_ embarrassingly parallel and requires quite an effort to code properly.

That is why BLAS/LAPACK is popular and there are few competing implementations.

vatican_banker··on Mathematica 14
I find impressive the breadth and depth of Wolfram's writing. I wish I could be that productive and write that much. Nevertheless, it's exhausting to read him because his texts are full of hubris.

This man needs to learn to edit himself.

vatican_banker··on A boy saw 17 doctors over 3 years for pain. ChatGPT found the right diagnosis
According to the article, it is difficult to diagnose babies/kids/toddlers because "[...] they can’t speak”.

The article also mentions that the kid didn't have the symptoms or physical manifestations of spina bifida, complicating the diagnosis.

This case reads as a genuinely difficult case to diagnose.

vatican_banker··on Companies in Japan opting for select offices to work in English
These companies are, most likely, targeting talent from China, South Korea, or South East Asia.

Top talent from, say, Vietnam already speaks good English and many would jump to the opportunity to work in Japan under these conditions.

vatican_banker··on Study links omega-3s to improved brain structure, cognition at midlife
> does not come with the problems of consuming fish

What kind of problems do mean? Environmental? Or health-related problems?

vatican_banker··on 911 Proxy Service Implodes After Disclosing Breach
I know at least two:

1. SEON: https://seon.io/

2.IPQS: https://www.ipqualityscore.com/

vatican_banker··on Top Performers Have a Superpower: Happiness
> the study doesn't measure happiness vs performance, it measures "how do you rate yourself" vs "the army's selection process for awarding medals".

The author explains that happiness is subjective:

> The behavioral science literature often refers to happiness as subjective well-being because the meaning of happiness varies in different contexts

...and then proceeds to delineate how psychology defines happiness:

1. a person’s own assessment of their satisfaction with life;

2. how much positive emotion [...] they experience;

3. and how little negative emotion [...] they experience

So yeah, self-rating is an important factor to measuring happiness.

vatican_banker··on I changed my mind about advertising
> Because advertising is manufactured demand.

I see advertising radically different.

People exhibit an spectrum of interest in products in the market, from “zero interest in buying” to “shut up and take my money”. Advertising works by convincing people close to the “shut up and take my money” part of the spectrum to actually buy.

Disclaimer: I’m not a marketeer.

vatican_banker··on Former Uber Chief Security Officer to Face Wire Fraud Charges
Yes, mentioned almost at the end:

> The separate guilty pleas entered by the hackers demonstrate that after Sullivan assisted in covering up the nature of the hack of Uber, the hackers were able to commit an additional intrusion at another corporate entity—Lynda.com—and attempt to ransom that data as well.

vatican_banker··on Polars: Fast DataFrame library for Rust and Python
In what way data.table trumps dplyr? Genuinely interested in knowing.

While data.table is faster than dplyr, data manipulations with data.table are difficult to read/understand/maintain.

dplyr also grew into a full-fledge list of libraries to work on data-related projects (the tidyverse). These libraries are _very_ well thought out and enables productivity with minimal learning curve [anecdotal]

vatican_banker··on Rows.com – Spreadsheet that supports external API integration and collaboration
Interesting... it doesn't seem to allow putting charts next to the spreadsheet though
vatican_banker··on Rows.com – Spreadsheet that supports external API integration and collaboration
grid.is looks more like a shiny app on steroids while Rows feel more like a hardcore spreadsheet. That being said, I haven't tried any of these.
vatican_banker··on Rows.com – Spreadsheet that supports external API integration and collaboration
Haven't heard from them before
vatican_banker··on Rows.com – Spreadsheet that supports external API integration and collaboration
The real power of this product is as a replacement of BI/Visualization tools. Imagine being able to connect Rows to a database and create "governed" sheets that look/work like dashboards with charts, tables, filters, etc.

It would be a killer product because most users are familiar with spreadsheets already. Many users end-up copying data from dashboards into spreadsheets (gsheet, excel, etc) so why not skipping the intermediaries and go straight to delivering a hybrid of dashboards and spreadsheets.

Lots of potential.

vatican_banker··on Voila – From notebooks to standalone web applications and dashboards
This is great. I've been looking for the equivalent of RShiny in the python world and never heard of streamlit before
vatican_banker··on Check If Email Exists
Financial/fintech companies use services like these for fraud-detection on account opening. While validating an email is by no means and exhaustive and conclusive signal to classify a fraud/genuine user, verifying the validity of new customers's email addresses is a big help.
vatican_banker··on History of the Nautilus loudspeaker
> It's a mistake to apply vanilla statistical thinking here. The 12 participants [...] were extremely skewed towards enthusiasts/professionals

It is still undetermined if having 12 highly-skilled professionals in the experiment is enough to have a conclusive experiment.

Also, this subject is so difficult to get right that the authors of the article themselves hedged by saying that experiment "does not support watertight conclusions".

vatican_banker··on History of the Nautilus loudspeaker
My honest and unscientific opinion is that the difference _is_ discernible but the listener needs to know what to hear for. Also, the reproduction quality is impacted by several factors like room, equipment, and recording quality (not just speaker quality).

[Anecdotal] One example of the difference between MP3 and lossless: the "image" [1] on 256kbps MP3s is worse compared to the the original uncompressed, lossless, versions (but the listening room must be appropriately prepared to reproduce a good image).

This is a highly subjective topic. IMO we'll never reach full agreement. Personally, I listen MP3 while on-the-go and lossless music at home.

Important to keep in mind the "size" of the experiment. Two interesting quotes from the article in c't magazine:

> twelve participants would be asked to come to Hanover.

> It's true that the data we collected does not support watertight conclusions, but they do provide interesting insights.

[1] https://en.wikipedia.org/wiki/Stereo_imaging

vatican_banker··on Practical advice for analysis of large, complex data sets (2016)
>If I had one piece of advice to give on the subject it would be PCA the crap out of everything and understand what the top components are doing

There are at least four issues with this advice (wrt doing data analysis):

1. How do you link your PCA components to the original data? Let's say you are tasked to find the main drivers of sales on a given city. You run PCA on the data and find two main components on the dataset. What do you do next? How do you make this information actionable?

2. How do you treat categorical variables? There are PCA methods for dealing with categorical variables but by the time you apply these methods plus the issues in 1) your data has lost all actionable meaning.

3. PCA is _very_ difficult to explain to business stakeholders. The more difficulty business stakeholders have to understand the analysis, the less they will use it.

4. Data-driven business stakeholders will favour clarity and simplicity over sophistication (somewhat linked to 3)

vatican_banker··on Kaspersky believes it found new CIA malware
There are several examples. This is one: https://www.theguardian.com/technology/2017/oct/26/kaspersky...
vatican_banker··on “Why We Sleep” Is Riddled with Scientific and Factual Errors
> ou shouldn't trust anybody other than yourself

This is not good advice.

I’d say find experts that you trust and be aware that these experts may commit mistakes as any other human.

In short: find the experts that make the least mistakes.

vatican_banker··on Collabora Office: The enterprise-ready edition of LibreOffice
My org uses Google Docs as office suite and I find it quite good and prefer it over any flavour of Microsoft Office. Nevertheless, there's a _huge_ and vocal push against Google Docs suite in favour of Microsoft Office from some internal stakeholders. None of the complaints from this group go beyond something to the tune of "I know Excel better".

As good as LibreOffice may be (I personally think it is _not_ better than Google docs) and as much as I like the romantic idea of open source office suite, I don't believe this will be successful/sustainable in the long run. The bulk of office suite users are not interested on using alternatives. These users want office running on their machine, period. I'd even venture to say that most users would rather have apps installed instead of Microsoft Office Online.

Page 1 of 2Next →