HNHacker News
TopNewBestAskShowJobs

mikeshi42

674 karma · joined July 8, 2014

Helping developers wrangle bugs and incidents: mike@hyperdx.io
submissionscomments
mikeshi42··on Show HN: I made a flight simulator, except you're just a passenger
3 overhead reading lights for a 2-2 configuration? On a trans pacific flight? Literally unplayable.

Would love to see this open source and extended!

mikeshi42··on Show HN: Make your logo extra bright on HDR screens
I’ve always wondered why it happened, assumed it was a rendering bug.

fwiw it seems like “reduce white point” overrides any of these effects - which is good.

mikeshi42··on Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
High CACs are supported by high LTVs, that is about the right ballpark for things like lawyers or even plumbers.
mikeshi42··on GPT-5.6
I've been using both and as far as I can tell with ccusage, the $ equivalent budgets are about the same between them now. This may have been true before anthropic doubled their quotas and openai 2x promo expired last month.
mikeshi42··on A successful Japanese trial of a ramjet engine designed for Mach‑5 aircraft
I have _yet_ to hit a time where TSA can make multiple hours disappear. Precheck w/ touchless ID lines are virtually empty at most airports, the actual security screen itself is quite fast given almost nothing needs to be removed from your bag these days. I still tend to arrive early, but I don't mind getting work done at the airport, especially at a lounge - though I've arrived very close to departure other times and still make it to the gate with plenty to spare.

On international returns, both Global Entry or MPC lines are virtually empty when I arrive (SFO)

The worst part is international arrivals in foreign countries, where immigration can soak up a lot of time, and you have no choice but to stand in line. Luckily I don't have to fly internationally too many times a year.

mikeshi42··on Appearing productive in the workplace
The irony is when I was sanity checking with ChatGPT - it caught the inaccuracies on its own.
mikeshi42··on Appearing productive in the workplace
While I agree with some of these observations - the research cited in the article really do not match the claims at all from what I can tell.

> An NBER study of support agents [2] found generative AI boosted novice productivity by about a third while barely helping experts. Harvard Business School researchers found the same pattern in consulting work [3].

The first work cited was a research study on GPT-3(!) from 2020. Which is a barely coherent model relative to today's SOTA.

The second HBS research study literally finds the opposite of what's claimed:

> we observed performance enhancements in the experimental task for both groups when leveraging GPT-4. Note that the top-half-skill performers also received a significant boost, although not as much as the bottom-half-skill performers.

Where bottom-half skilled participants with AI outperformed top-half skilled participants without AI. (And top-half skilled participants gained another 11% improvement when pared with AI). Again, GPT-4 model intelligence (3 years ago) is a far cry from frontier models today.

mikeshi42··on LinkedIn is searching your browser extensions
I’ve seen fake accounts created by bad actors attempting to pose as others for gaining remote employment. It’s possible that is what was happening, and the takedown was from LI taking down the profile from the bad actor.

Other times they would just link to real LinkedIn profiles, but the LinkedIn profile will say that they’re not actively looking and are a victim of id fraud basically.

It’s been a huge issue spotting candidates falsifying information since remote work took off unfortunately. They payout is if they can get at least 1 or 2 paychecks before being found out, they’ve made a good profit.

mikeshi42··on LLM Pareto Frontier
I'm actually curious on the evolution of models as well over time - I created a "created at" group by option to try to broadly visualize how models over time (half calendar years) drift in terms of pareto efficiency. There's a clear trend going up which is good.

I hope to build some better visualizations down the line on model evolution instead of just hacking it onto the current scatterplot.

mikeshi42··on The current state of LLM-driven development
Yes, though that just means the probability of success is a function of not only user input but also the model version.

Slot machines on the other hand are truly random and success is luck based with no priors (the legal ones in the US anyways)

mikeshi42··on The current state of LLM-driven development
There's plenty of evidence that good prompts (prompt engineering, tuning) can result in better outputs.

Improving LLM output through better inputs is neither an illusion, nor as easy as learning how to google (entire companies are being built around improving llm outputs and measuring that improvement)

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
This is really helpful feedback, a few notes:

- Reorder columns: You should be able to do so by modifying the order of the SELECT statement at the top.

- Wrapping: totally agree - this is something we're adding in, it's on the near term todos.

- Pin fields: you're totally right, you can pin values but that doesn't prioritize the filter in general. I'm getting this one in the queue.

- Time picker: I hear you with Google Cloud's time picker, they did indeed make a really nice one. Though I'm curious to hear more about the breakdown, is it wanting to customize how that chart is grouped beyond log level?

- Autocomplete: We had a regression if you're on v2.0.0 which is what I suspect you're hitting. If you're on v2.0.1 lucene should have field value autocomplete. We don't have it yet for SQL.

- Live View: Can you clarify what is scrolling? Or is this about disabling live?

- Chart View: I just tried really quick and it seems okay to me, do you mind sharing more details?

- Sidebar: Is this related to field popularity to pin fields or something else?

This feedback is all super helpful - you can probably tell we're still early in building out the perfect experience. I'd love it if we could dive in deeper either on our discord (https://hyperdx.io/discord) or email: michael.shi@clickhouse.com. Either way this has been exceptionally helpful :)

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
We think Grafana is still a great tool and many teams are heavily invested in the Grafana ecosystem. We'll continue to invest in Grafana support via ClickHouse's official Grafana plugin and that won't be changing at all with this release.

However, there's a bit of a fundamental difference in the user experience we're targeting. Grafana has really excelled at traditional monitoring dashboards, low cardinality monitoring workflows.

ClickHouse unlocks a newer paradigm of high cardinality, high performance observability. It enables a new set of workflows/UX that allows engineers to query novel problems quickly as opposed to working off of static dashboards. That's really a big focus of ours, so you'll see we do exploration/search/syntax/UI layout is quite different from Grafana due to this.

At this point it isn't even an original realization of ours. Just as an example, Shopify built a complete custom app (only keeping the auth part of Grafana) while migrating to ClickHouse for similar reasons.

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
Ah sorry I missed that part of the question, yes MongoDB and ClickHouse are the two stateful services. We'll be looking to see if we can offer some mode to simplify it down to just ClickHouse but that'll take a bit more work.
mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
For getting started, we have a really easy to use helm chart that I'd recommend checking out first: https://clickhouse.com/docs/use-cases/observability/clicksta...
mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
I've written up a detailed answer in a thread below earlier :) https://news.ycombinator.com/item?id=44196484
mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
ClickStack is currently just open source - so there's no cloud or a fully hosted offering yet! (Of course you can always pair ClickStack with ClickHouse Cloud to have your ClickHouse hosted for you).

But in this case there's probably no reason for you :) These improvements will come to our cloud offering of course as we work on rolling out upgrades from HyperDX v1 to v2 in cloud.

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
It's actually not clear to me that Vector is any simpler than OTel imo (VRL is way more complicated than OTTL for instance). You can also use otel collector builder (ocb) to build a slimmed binary.

My take is that OTel is overall the best investment, it's widely supported across the board by many companies and other vendors. It's also constantly being improved with interesting ideas like otel-arrow which will make it even more performant (and columnar friendly!)

We'll also continue invest in the OTel ecosystem ourselves in making it easier and easier to get started :)

That being said, I'm not saying that OTel collector is always the right choice, we want to meet users where they are. Some users have data that gets piped into S3 files and we ingest off of a S3 bucket just due to how they've collected data, some use Vector due to its flexibility with VRL, focus on logs, or specific integrations it provides out of the box. So the answer is always - it depends :) but I do like OTel and think the future is bright.

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
Totally agree - you use an observability tool because it answers your questions quickly, not just return searches quickly.

Beyond raw performance and cost effectiveness, which is quite important at scale, we work a lot on making sure the application layer itself is intuitive to use. You can always play around with what ours looks like at play.hyperdx.io :)

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
There's a version that we call local mode which is intended for engineers using it as part of their local debugging workflow: https://clickhouse.com/docs/use-cases/observability/clicksta...

Otherwise yes you can authenticate against the other versions with a email/password (really the email doesn't do anything in the open source distribution, just a user identifier but we keep there to be consistent)

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
This is good feedback to make things more clear :) HyperDX is part of ClickStack, so ClickStack = { HyperDX, ClickHouse, OTel }. This is the stack we recommend that will deploy in seconds and _just work_, and can scale up to PB+ and beyond as well with some additional effort (more than a few seconds unfortunately, but one day...)

HyperDX v2, the version that is now stable and shipped in ClickStack, focuses more on the querying layer. It lets users have more customization around ClickHouse (virtually any schema, any deployment).

Optionally, users can leverage other ways of getting data into ClickHouse like Vector, S3, etc. but still use HyperDX v2 on top. Previously in HyperDX v1 you _had_ to use OTel and our ingestion pipeline and our schemas. This is no longer true in v2.

Let me know if this explanation helps

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
That'd be awesome! Ferret has been on my radar for a while now :) If you want to chat with us on Discord: https://hyperdx.io/discord
mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
Nope! We're virtually schema agnostic, you can map your custom schema to observability concepts (ex. the SQL expression for TraceID, either a column or a full function/expression will work).

We don't have any lock in to our ingestion pipeline or schema. Of course we optimize a lot for the OTel path, but it works perfectly fine without it too.

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
In theory you should be able to try using FerretDB for example.

We have this on the medium term roadmap to investigate proper support for a compatibilty layer such as ferret or more likely just using ClickHouse itself as the operational data store.

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
Luckily ClickHouse and serious throughput are pretty synonymous. Internally we're at 100+PB of telemetry stored in our own monitoring system.

Vector supports directly writing into ClickHouse - several companies use this at scale (iirc Anthropic does exactly this, they spoke about this recently at our user conference).

Please give it a try and let us know how it goes! Happy to help :)

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
Echoing the comment below, I guess one obvious thing is that we are a team at ClickHouse and an official first-party product on top. That translates into:

- We're flexible on top of any ClickHouse instance, you can use virtually any schema in ClickHouse and things will still work. Custom schemas are pretty important for either tuned high performance or once you're at a scale like Anthropic. This makes it also incredibly easy to get started (especially if you already have data in ClickHouse). - The above also means you don't need to buy into OTel. I love OTel but some companies choose to use Vector, Cribl, S3, a custom writing script, etc for good reasons. All of that is supported natively due to the various ClickHouse integrations, and naturally means you can use ClickStack/HyperDX in that scenario as well. - We also have some cool tools around wrangling telemetry at scale, from Event Deltas (high cardinality correlation between slow spans and normal spans to root cause issues) to Event Patterns (clustering similar logs or spans together automatically with ML) - all of these help users dive into their data in easier ways than just searching & charting. - We also have session replay capability - to truly unify everything from click to infra metrics.

We're built to work at the 100PB+ scale we run internally here for monitoring ClickHouse Cloud, but flexible enough to pin point specific user issues that get brought up once in a support case in an end-to-end manner.

There's probably a lot more I'm missing. Ultimately from a product philosophy standpoint, we aren't big believers in the "3 pillars" concept, which tends to manifest as 3 silos/tabs for "logs", "metrics", "traces" (this isn't just Signoz - but across the industry). I'm a big believer that we're building tools to unify and centralize signals/clues in one place and giving the right datapoint at the right time to the engineer. During an incident I just think about what's the next clue I can get to root cause an issue, not if I'm in the logging product or the tracing product.

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
Wow this is an incredible throwback! Can't believe your memory is this good. It's quite funny and I totally agree - I met the Gradio founders in an accelerator (when they were just getting started) after we shut down ModelDepot - and they of course ended up getting acquired into Hugging Face. It's funny how things end up sometimes :)
mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
First off, always really excited to hear from our production users - glad to hear you're getting good value out of the platform!

HyperDX isn't being deprecated, you can probably see on the marketing page it's still really prominently featured as an integral part of the stack - so nothing changing there.

We do of course want to get users onto HyperDX v2 and the overall ClickStack pattern. This doesn't mean HyperDX is going away by any means - just that HyperDX is focused a lot more on the end-user experience, and we get to leverage the flexibility, learnings and performance of a more exposed ClickHouse-powered core which is the intent of ClickStack. On the engineering side, we're working on making sure it's a smooth path for both open source and cloud.

side note: weird I thought I replied to this one already but I've been dealing with spotty wifi today :)

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
Great questions!

OTel Metrics: I get it, it's specified as almost a superset of everyone's favorite metric standards with config for push/pull, monotonic vs delta, exponential/"native" histograms, etc. I have my preferences as well which would be a subset of the standard but I get why a unifying standard needed to be flexible.

Statsd: The great thing about the OTel collector is that it allows ingesting a variety of different data formats, so you can take in statsd and output OTel or write directly to ClickHouse: https://github.com/open-telemetry/opentelemetry-collector-co...

We correlate across trace/span id as well as resource attributes. The correlation across logs/traces with span/trace id is a pretty well worn path across our product. Metrics to the rest is natively done via resource attributes and we primarily expose correlation for K8s-based workloads with more to come. We don't do exemplars _yet_ to solve the more generic correlation case for metrics (though I don't think statsd can transmit exemplars)

Elixir: We try to do our best to support wherever our users are, the OTel SDK and ours have continued to change in parallel over time - we'll want to likely re-evaluate if we should start pointing towards the base OTel SDK for Elixir. We've been pretty early on the OTel SDK side across the board so things continue to evolve, for example our Deno OTel integration came out I think over a year before Deno officially launched one with native HyperDX documentation <3

Notebooks: Yes, it should land in an experimental state shortly, stay tuned :) There's a lot of exciting workflows we're looking to unlock with notebooks as well. If you have any thoughts in this direction, please let me know. I'd love to get more user input ahead of the first release.

mikeshi42··on Show HN: ClickStack – Open-source Datadog alternative by ClickHouse and HyperDX
I agree - rotel seems like a really good fit for a lightweight lambda integration for OTel, it of course should work already since we stand up an OTel ingest endpoint so it should be seamless to send data over! (Kind of the beauty of OTel of course)

I've also been in touch with Mike & Ray for a bit, who've told me they've added ClickHouse support recently which makes the story even better :)

Page 1 of 8Next →